Live Feed
OpenRouter Ships Its Own 100B Model — Elephant Alpha Is Free on Day One
OpenRouter, which routes API traffic to 200+ third-party models, has released its own 100B-parameter model at $0 per token. The move signals a strategic shift from pure aggregator to model producer — and raises questions about what Elephant's free pricing is actually optimising for.
US Federal Agencies Are Using AI to Write Regulations — DOT Names Google Gemini Its Rulemaking Tool
The Department of Transportation is using Google Gemini to draft proposed federal rules, calling it the "point of the spear" for AI in government. DOGE separately proposed using an LLM to rescind half of all federal regulations. AI has entered the rulemaking chain — and the legal implications are unresolved.
CoreWeave Closes $8.5B Investment-Grade GPU Loan — First A3-Rated AI Infrastructure Debt
CoreWeave secured an $8.5 billion delayed draw term loan on March 31, rated A3 by Moody's — the first investment-grade financing backed by GPU clusters and customer contracts. The deal, anchored by a $19B Meta backlog, marks the formal arrival of AI compute as an institutional asset class.
Musk Sets June Deadline for Grok to Match Claude Opus 4.6 — After Admitting xAI Was "Not Built Right"
Elon Musk said Grok will be close to Claude Opus 4.6 by May and match it by June. The admission follows a March acknowledgement that xAI is undergoing a full architectural rebuild — and a public concession that it currently trails OpenAI, Anthropic, and Google by roughly seven months.
New York RAISE Act Is Finalised: 72-Hour Incident Window, $1M Penalties, Effective January 2027
New York signed the final version of its frontier AI safety law on March 27, giving large developers nine months to comply. The 72-hour critical incident disclosure window is stricter than California's 15-day equivalent — creating a new compliance bar for any model operating in the state.
The 36x Reasoning Tax: DeepSeek R1 vs o3-pro in the 2026 API Economy
DeepSeek R1 charges $0.55/M input tokens. OpenAI's o3-pro charges $20/M. The gap is 36x on input and 36x on output — while R1 posts within 2 points of o3's graduate-level science score. This is the cost calculus defining production reasoning in 2026.
Anthropic Launches AI Vulnerability Scanner with $104M Coalition — Amazon, Apple, Google, Microsoft, NVIDIA Back Project Glasswing
Claude Code Security entered limited preview on April 11 for Enterprise and Team accounts. It uses AI reasoning to trace data flows and identify vulnerabilities missed by static analysis. The launch is part of Project Glasswing, a $104M industry initiative backed by five major tech companies and 40+ critical infrastructure organisations.
KellyBench: Every Frontier AI Model Loses Money Betting the Premier League
General Reasoning releases KellyBench, a benchmark that forces LLM agents to manage a £100,000 bankroll across a full Premier League season. Every tested model loses money. Claude Opus 4.6 finished best at -11% ROI. Several models went to £0.
NVIDIA Vera Rubin: 50 PFLOPS FP4, 3nm Chiplets, and a New GPU Built Purely for Long-Context Inference
NVIDIA's Vera Rubin architecture lands H2 2026 with 336 billion transistors on TSMC 3nm — a 1.6x density jump from Blackwell — and 50 PFLOPS FP4 per GPU. A new product, the Rubin CPX, targets long-context inference with no Blackwell equivalent.
Anthropic Run-Rate Revenue Hits $30B, Company Explores Custom AI Chips
Anthropic's annualised revenue has surpassed $30 billion — up from $9 billion at end of 2025 — driven by accelerating Claude demand. The company is now exploring designing its own AI chips, mirroring moves by Meta and OpenAI.
ByteDance Launches Seeduplex: First Production-Scale Full-Duplex Voice AI
ByteDance deploys Seeduplex in Doubao, the first voice AI model that listens and responds simultaneously at scale. False interruption rates drop 50% versus the prior half-duplex system. End-of-turn detection is 250ms faster.
Image Generation API Shootout 2026: GPT Image 1.5 Leads Text, Flux 2 Pro Leads Price, Imagen 4 Leads Photo
A comprehensive April 2026 analysis of every major image generation API finds three distinct category leaders — with per-image costs ranging from $0.02 to $0.12 depending on provider, resolution, and quality tier. DALL-E 3 remains most integrated but has lost the quality crown.
Intel and Google Sign Multiyear AI Infrastructure Pact — CPUs Fight Back Against the GPU Monoculture
Intel and Google formalized a multiyear deal on April 9 to deepen Xeon deployment across Google Cloud and co-develop custom Infrastructure Processing Units (IPUs). No financial terms were disclosed, but the partnership signals a deliberate counter-narrative: AI infrastructure needs more than GPUs.
Meta Locks In $21B CoreWeave AI Cloud Through 2032, Anchors Vera Rubin Rollout
CoreWeave and Meta have expanded their existing AI infrastructure agreement to $21 billion through December 2032, with capacity spanning multiple locations and including early Vera Rubin GPU cluster deployments.
OpenAI Launches $100/Month Codex Pro 5X Plan After Token-Based Metering Guts Business Tier
OpenAI introduced a $100/month Codex Pro 5X plan on April 9 delivering 200-1,000 GPT-5.4 messages per 5-hour window. A simultaneous switch to token-based metering has left Business-tier developers reporting effective throughput drops of 2.5-5x, forcing mid-session model swaps.