GPT-56T 861 —
MUSE-SPK 835 -0.7%
GPT-56SC 828 -5.2%
QWEN-38X 824 —
CL-OP55X 822 —
GROK-46H 822 -5%
GPT-6A 820 —
GLM-5 784 -8.4%
CL-FAB5H 743 -5.6%
KIMI-K3X 742 -8.4%
CL-OP5H 720 -5.8%
CL-OP5X 709 -18%
CL-OP46H 698 -5.9%
CL-OP47H 690 -5.9%
GEM-38FH 677 +0.1%
GEM-37FH 657 -24%
GPT-56S 622 —
CL-OP47 582 -0.7%
GPT-55H 582 —
INKL 531 —
GEM-31P 513 —
GEM-3P 499 —
CL-OP46 496 -0.2%
CL-OP48 490 —
GPT-56T 861 —
MUSE-SPK 835 -0.7%
GPT-56SC 828 -5.2%
QWEN-38X 824 —
CL-OP55X 822 —
GROK-46H 822 -5%
GPT-6A 820 —
GLM-5 784 -8.4%
CL-FAB5H 743 -5.6%
KIMI-K3X 742 -8.4%
CL-OP5H 720 -5.8%
CL-OP5X 709 -18%
CL-OP46H 698 -5.9%
CL-OP47H 690 -5.9%
GEM-38FH 677 +0.1%
GEM-37FH 657 -24%
GPT-56S 622 —
CL-OP47 582 -0.7%
GPT-55H 582 —
INKL 531 —
GEM-31P 513 —
GEM-3P 499 —
CL-OP46 496 -0.2%
CL-OP48 490 —

Live Feed

3mo ago policy

Uber Caps Employee AI Spending at $1,500/Month After Budget Collapses in Four Months

Uber has set a hard $1,500/month per-employee ceiling on AI tools including Claude Code, after exhausting its entire 2026 AI budget by April. The cap is the clearest enterprise signal yet that frontier coding agent costs are structurally incompatible with uncapped deployment.

3mo ago research

Biohub's ESMFold2 Outperforms AlphaFold on Antibodies and Designs Working Binders Against 5 Cancer Targets

Chan Zuckerberg Biohub released a three-part protein world model on May 27: ESMC (trained on 2.8 billion protein sequences), ESMFold2 (beats AlphaFold on antibody-antigen complexes, no multiple sequence alignments required), and ESM Atlas (1.1 billion predicted structures, 6.8 billion searchable proteins). All three released under MIT license. Binders designed by the system have been validated in laboratory experiments against cancer and immunology targets.

3mo ago policy

130 Signatories Back Leiden Declaration: AI Is Threatening the Integrity of Mathematical Proof

The International Mathematical Union and Fields Medal laureate Peter Scholze endorse an 11-page declaration from 16 mathematicians at 15 universities, published June 2. It names five structural threats AI poses to mathematics and issues 23 recommendations — including mandatory disclosure, human accountability for correctness, and investment in public AI research.

3mo ago benchmark

All Three Memory Giants Cross $1 Trillion in One Month — HBM Supply Sold Out, Shortages Through 2030

SK Hynix and Micron both crossed $1 trillion in market cap within 24 hours of each other in late May, joining Samsung. The catalyst: high-bandwidth memory demand has outpaced supply, with every chip either company can produce through 2026 already spoken for.

3mo ago research

AI Hovers Near Chance at Predicting Scientific Discoveries — Paper Tests 4,760 Events and Finds a Hard Ceiling on Foresight

A new paper tests frontier AI models across 4,760 scientific events and finds a large gap between recognition and prediction. Models are strong at identifying plausible research paths when answers are nearby, but hover near chance when asked whether a discovery will actually happen, when it will arrive, or what method will make it work.

3mo ago model

Microsoft MAI-Thinking-1 Hits 97% on AIME — First Reasoning Model Built Without Distillation, Matches Opus 4.6 on SWE-Bench Pro

Microsoft shipped its first in-house reasoning model at Build 2026: MAI-Thinking-1 is a 35B-active / 1T-parameter MoE trained from scratch on clean data with no distillation from third-party models. It scores 97.0% on AIME 2025, matches Claude Opus 4.6 on SWE-Bench Pro, and is in private preview on Microsoft Foundry.

3mo ago research

Stanford Blind Study: AI Wins 75% of Law Professor Comparisons — and Gets Flagged as Harmful 3.4x Less Often

A Stanford Law study by Professor Julian Nyarko tested 16 law professors in 2,885 blind head-to-head comparisons of AI versus peer answers to contract-law student questions. AI won 75% of matchups. Professors rated AI responses as pedagogically harmful 3.5% of the time, versus 12% for answers from fellow instructors.

3mo ago policy

Trump Signs Downsized AI Order: 30-Day Voluntary Review Window, Cybersecurity Focus, No Hard Mandates

After canceling a 90-day pre-release review requirement hours before signing last month, the White House has signed a scaled-back AI executive order: a 30-day voluntary government review window for powerful new models, with an explicit cybersecurity framing and no binding compliance mechanism.

3mo ago release

Microsoft MAI-Code-1-Flash: 51.2% SWE-Bench Pro, +16 Points Over Haiku 4.5, 60% Fewer Tokens

Microsoft's Superintelligence team ships MAI-Code-1-Flash, a new coding model built end-to-end by Microsoft, rolling out to GitHub Copilot users in VS Code. It posts 51.2% on SWE-Bench Pro versus Claude Haiku 4.5's 35.2% — and solves harder problems with up to 60% fewer tokens.

3mo ago policy

Glasswing Grows to 200 Total Partners: 150 New Orgs in 15 Countries, Power Grid and Healthcare Added

Anthropic is extending Project Glasswing from its initial 50-partner cohort to roughly 150 new organizations across 15+ countries. New sectors include power, water, healthcare, and hardware vendors — each chosen on the basis that a successful attack on their codebase could affect more than 100 million people.

3mo ago release

JetBrains Open-Sources Mellum2: 12B MoE at 2x Speed for Agentic Sub-Task Routing

JetBrains released Mellum2 under Apache 2.0: a 12B MoE model with 2.5B active parameters, 2x faster inference than comparable open models, and a design brief aimed squarely at agentic infrastructure — routing, RAG, sub-agents, and private deployment.

3mo ago release

NVIDIA Cosmos 3: First Open Physical AI Omnimodel Ships at Computex — 20T Tokens, Robots Get a Physics Brain

NVIDIA released Cosmos 3 at GTC Taipei: the first open model to unify vision reasoning, world generation, and robot action prediction in one architecture. Nano runs on a workstation. Super targets datacenters. #1 across 7+ open-model robotics benchmarks.

3mo ago release

MiniMax M3 Launches: 59% SWE-Bench Pro, 70% Computer Use, 15x Faster Long-Context Than M2

MiniMax's M3 clears GPT-5.5 and Gemini 3.1 Pro on SWE-Bench Pro, hits 70.06% on OSWorld for computer use, and delivers 15x decoding speedup at 1M tokens via a new sparse attention architecture. Open weights within 10 days.

3mo ago policy

Florida Files the First State Lawsuit Against OpenAI and Sam Altman — FSU Shooting Chat Logs Cited

Florida Attorney General James Uthmeier sued OpenAI and CEO Sam Altman on June 1 under the state's Deceptive and Unfair Trade Practices Act, alleging ChatGPT aided a mass shooter, drove users to suicide, and addicted minors. It is the first state-level lawsuit against OpenAI in the US.

3mo ago funding

Groq Raises $650M to Rebuild as Inference Neocloud After Nvidia Took Its IP and Founders

Six months after a $20B Nvidia licensing deal stripped Groq of its architecture, founders, and engineering leadership, the company is raising $650M from existing investors to build an AI inference cloud. The window to survive is 18 to 24 months.