GPT-56T 861 —
MUSE-SPK 837 +0.2%
GPT-56SC 790 -4.6%
GLM-5 781 -0.4%
CL-OP55X 780 -5.1%
GROK-46H 780 -5.1%
QWEN-38X 748 -9.2%
GPT-6A 743 -9.4%
KIMI-K3X 742 —
CL-FAB5H 698 -6.1%
CL-OP5H 675 -6.2%
GEM-38FH 672 -0.7%
CL-OP5X 670 -5.5%
CL-OP55H 668 —
CL-OP46H 657 -5.9%
CL-OP47H 648 -6.1%
GPT-56S 618 -0.6%
GEM-37FH 610 -7.2%
GEM-36FH 593 —
CL-OP48H 588 —
CL-OP47 581 -0.2%
GEM-35FH 580 —
GPT-55H 541 -7%
INKL 531 —
GEM-31P 512 -0.2%
CL-OP46 498 +0.4%
GEM-3P 498 -0.2%
CL-OP48 492 +0.4%
GPT-52 464 —
GPT-55 423 —
GPT-56T 861 —
MUSE-SPK 837 +0.2%
GPT-56SC 790 -4.6%
GLM-5 781 -0.4%
CL-OP55X 780 -5.1%
GROK-46H 780 -5.1%
QWEN-38X 748 -9.2%
GPT-6A 743 -9.4%
KIMI-K3X 742 —
CL-FAB5H 698 -6.1%
CL-OP5H 675 -6.2%
GEM-38FH 672 -0.7%
CL-OP5X 670 -5.5%
CL-OP55H 668 —
CL-OP46H 657 -5.9%
CL-OP47H 648 -6.1%
GPT-56S 618 -0.6%
GEM-37FH 610 -7.2%
GEM-36FH 593 —
CL-OP48H 588 —
CL-OP47 581 -0.2%
GEM-35FH 580 —
GPT-55H 541 -7%
INKL 531 —
GEM-31P 512 -0.2%
CL-OP46 498 +0.4%
GEM-3P 498 -0.2%
CL-OP48 492 +0.4%
GPT-52 464 —
GPT-55 423 —

Live Feed

5mo ago research

NVIDIA Ships Ising: Open AI Models for Quantum Error Correction, 2.5x Faster Than pyMatching

NVIDIA launched Ising, the first open-source AI model family purpose-built for quantum computing infrastructure. Two models — a vision-language calibrator and a 3D CNN error decoder — tackle the two engineering bottlenecks blocking practical quantum hardware. IONQ surged 50%, QUBT 30% on the announcement.

5mo ago research

Cross-Datacenter LLM Inference Is Now Viable: Prefill-as-a-Service Delivers 54% Throughput Gain

A new paper from researchers running a 1-trillion-parameter hybrid model shows that splitting LLM inference across datacenters — routing long prompts to a remote prefill cluster while keeping decode local — delivers 54% higher throughput over local-only serving and 32% over naive heterogeneous setups. The enabling factor is smaller KV cache in hybrid-attention architectures.

5mo ago policy

Tinder and Zoom Adopt Iris Scanning as AI Makes Faces Untrustworthy

World — formerly Worldcoin — has signed Tinder and Zoom as its first major platform partners for World ID, iris-based proof of humanity. The move signals that biometric personhood verification is becoming a practical layer on apps flooded with AI-generated fakes.

5mo ago benchmark

First-Ever Three-Way Tie at Intelligence Index 57: All Three Frontier Labs Now Equal

Artificial Analysis has called a tie at the top of its Intelligence Index for the first time. Claude Opus 4.7 (57.3), Gemini 3.1 Pro Preview (57.2), and GPT-5.4 (56.8) all round to 57 within a 1-point confidence interval. The composite ranking is dead. Lab differentiation is the new story.

5mo ago special-report

Special Report: The AI Index 2026 — Signals From a Frontier That Refuses to Slow Down

The ninth AI Index puts hard numbers on a year the frontier kept accelerating, the US–China gap effectively closed, private capital bent toward the United States, and the hardware stack concentrated around a single foundry. A data-forward read of where the model race actually is.

5mo ago release

Anthropic Puts Claude Inside Microsoft Word With Native Tracked Changes — Legal Review Is the Opening Move

Anthropic has released a beta add-in placing Claude directly inside Microsoft Word. Every AI-generated edit surfaces as a native tracked change — not a sidebar suggestion. Legal contract review is the lead use case, signalling Anthropic is going after the enterprise document workflow that anchors Word's commercial moat.

5mo ago release

llama.cpp Merges Speculative Checkpointing: 40% Less VRAM, 20% More Throughput for Consumer 70B Inference

Georgi Gerganov merged speculative checkpointing into llama.cpp on April 18, cutting VRAM usage by up to 40% during batched operations and boosting token throughput by 15–20% on consumer hardware. The update makes extended-context 70B inference viable on setups that previously ran out of memory.

5mo ago benchmark

Opus 4.7 Opens a 79-Point Agentic Elo Lead Over GPT-5.4 — and Runs Cheaper Per Task Than Its Predecessor

Claude Opus 4.7 scores 1,753 on GDPval-AA, 79 Elo points ahead of Sonnet 4.6 and GPT-5.4 (1,674) and 134 ahead of Opus 4.6 (1,619). Despite identical list pricing at $5/$25 per million tokens, effective per-task cost is lower than Opus 4.6 Adaptive Reasoning — driven by reduced output token usage in Opus 4.7's new tokenizer.

5mo ago model

Opus 4.7's Hidden Cost: New Tokenizer Adds Up to 35% More Tokens Per Prompt

Anthropic held the line at $5/$25 per million tokens for Claude Opus 4.7, but confirmed the new model uses up to 35% more tokens for the same input. Developer reports also flag a collapse in long-context retrieval performance — from 78.3% to 32.2% on MRCR.

5mo ago release

xAI Opens Grok Voice APIs: $0.10/hr Transcription, $4.20/M Characters Speech Synthesis

xAI has launched standalone Grok Speech-to-Text and Text-to-Speech APIs built on the same production stack powering Grok mobile, Tesla vehicles, and Starlink customer support. The release positions xAI directly against ElevenLabs, Deepgram, and AssemblyAI with sharp pricing and a proven infra pedigree.

5mo ago release

Anthropic Launches Claude Design: Conversational Prototyping That Sent Figma Down 7%

Anthropic shipped Claude Design in research preview on April 17, putting prototypes, wireframes, pitch decks, and marketing assets behind a text prompt. Figma shares dropped 6.84% on the announcement. The tool runs on Opus 4.7, exports to Canva/PDF/PPTX/HTML, and hands off directly to Claude Code.

5mo ago funding

Cerebras Files for Nasdaq IPO With $20B OpenAI Contract and $510M in Revenue

AI chipmaker Cerebras Systems disclosed its S-1 on April 17, revealing $510M in 2025 revenue, a $20B multi-year OpenAI deployment deal, and an AWS data center integration. The listing breaks an 18-month freeze on AI hardware IPOs.

5mo ago release

Cloudflare Ships Five Agent Products in Three Days — Project Think, Mesh, and the Bid to Own Agent Infrastructure

Cloudflare ran "Agents Week" April 15–17, launching Project Think (a batteries-included Agents SDK with durable execution and sub-agents), Cloudflare Mesh (private networking for agent fleets), Agent Lee (dashboard AI assistant), Redirects for AI Training, and an Agent Readiness Score. It is the most concentrated infrastructure land-grab since the cloud wars.

5mo ago funding

Meta to Cut 8,000 Jobs Starting May 20 — The Largest AI-Driven Headcount-to-Compute Swap in Tech History

Meta will begin laying off roughly 10% of its global workforce on May 20, 2026, with additional rounds expected later in the year. The capital is being redirected into chips, data centres, and model training — making this the single largest AI-driven reallocation of labour to compute in tech history.

5mo ago model

MiniMax Open-Sources M2.7 at 1495 GDPval-AA ELO and 57% Terminal Bench — Self-Evolving Agent Hits Top-5

MiniMax has released M2.7 under an open-source license, posting 1495 ELO on GDPval-AA (highest at release), 56.22% on SWE-Pro, and 57.0% on Terminal Bench 2. A Chinese lab has put a freely self-hostable model inside the global top tier on agentic capability benchmarks.