Live Feed
Uber Caps Employee AI Spending at $1,500/Month After Budget Collapses in Four Months
Uber has set a hard $1,500/month per-employee ceiling on AI tools including Claude Code, after exhausting its entire 2026 AI budget by April. The cap is the clearest enterprise signal yet that frontier coding agent costs are structurally incompatible with uncapped deployment.
Biohub's ESMFold2 Outperforms AlphaFold on Antibodies and Designs Working Binders Against 5 Cancer Targets
Chan Zuckerberg Biohub released a three-part protein world model on May 27: ESMC (trained on 2.8 billion protein sequences), ESMFold2 (beats AlphaFold on antibody-antigen complexes, no multiple sequence alignments required), and ESM Atlas (1.1 billion predicted structures, 6.8 billion searchable proteins). All three released under MIT license. Binders designed by the system have been validated in laboratory experiments against cancer and immunology targets.
130 Signatories Back Leiden Declaration: AI Is Threatening the Integrity of Mathematical Proof
The International Mathematical Union and Fields Medal laureate Peter Scholze endorse an 11-page declaration from 16 mathematicians at 15 universities, published June 2. It names five structural threats AI poses to mathematics and issues 23 recommendations — including mandatory disclosure, human accountability for correctness, and investment in public AI research.
All Three Memory Giants Cross $1 Trillion in One Month — HBM Supply Sold Out, Shortages Through 2030
SK Hynix and Micron both crossed $1 trillion in market cap within 24 hours of each other in late May, joining Samsung. The catalyst: high-bandwidth memory demand has outpaced supply, with every chip either company can produce through 2026 already spoken for.
AI Hovers Near Chance at Predicting Scientific Discoveries — Paper Tests 4,760 Events and Finds a Hard Ceiling on Foresight
A new paper tests frontier AI models across 4,760 scientific events and finds a large gap between recognition and prediction. Models are strong at identifying plausible research paths when answers are nearby, but hover near chance when asked whether a discovery will actually happen, when it will arrive, or what method will make it work.
Microsoft MAI-Thinking-1 Hits 97% on AIME — First Reasoning Model Built Without Distillation, Matches Opus 4.6 on SWE-Bench Pro
Microsoft shipped its first in-house reasoning model at Build 2026: MAI-Thinking-1 is a 35B-active / 1T-parameter MoE trained from scratch on clean data with no distillation from third-party models. It scores 97.0% on AIME 2025, matches Claude Opus 4.6 on SWE-Bench Pro, and is in private preview on Microsoft Foundry.
Stanford Blind Study: AI Wins 75% of Law Professor Comparisons — and Gets Flagged as Harmful 3.4x Less Often
A Stanford Law study by Professor Julian Nyarko tested 16 law professors in 2,885 blind head-to-head comparisons of AI versus peer answers to contract-law student questions. AI won 75% of matchups. Professors rated AI responses as pedagogically harmful 3.5% of the time, versus 12% for answers from fellow instructors.
Trump Signs Downsized AI Order: 30-Day Voluntary Review Window, Cybersecurity Focus, No Hard Mandates
After canceling a 90-day pre-release review requirement hours before signing last month, the White House has signed a scaled-back AI executive order: a 30-day voluntary government review window for powerful new models, with an explicit cybersecurity framing and no binding compliance mechanism.
Microsoft MAI-Code-1-Flash: 51.2% SWE-Bench Pro, +16 Points Over Haiku 4.5, 60% Fewer Tokens
Microsoft's Superintelligence team ships MAI-Code-1-Flash, a new coding model built end-to-end by Microsoft, rolling out to GitHub Copilot users in VS Code. It posts 51.2% on SWE-Bench Pro versus Claude Haiku 4.5's 35.2% — and solves harder problems with up to 60% fewer tokens.
Glasswing Grows to 200 Total Partners: 150 New Orgs in 15 Countries, Power Grid and Healthcare Added
Anthropic is extending Project Glasswing from its initial 50-partner cohort to roughly 150 new organizations across 15+ countries. New sectors include power, water, healthcare, and hardware vendors — each chosen on the basis that a successful attack on their codebase could affect more than 100 million people.
JetBrains Open-Sources Mellum2: 12B MoE at 2x Speed for Agentic Sub-Task Routing
JetBrains released Mellum2 under Apache 2.0: a 12B MoE model with 2.5B active parameters, 2x faster inference than comparable open models, and a design brief aimed squarely at agentic infrastructure — routing, RAG, sub-agents, and private deployment.
NVIDIA Cosmos 3: First Open Physical AI Omnimodel Ships at Computex — 20T Tokens, Robots Get a Physics Brain
NVIDIA released Cosmos 3 at GTC Taipei: the first open model to unify vision reasoning, world generation, and robot action prediction in one architecture. Nano runs on a workstation. Super targets datacenters. #1 across 7+ open-model robotics benchmarks.
MiniMax M3 Launches: 59% SWE-Bench Pro, 70% Computer Use, 15x Faster Long-Context Than M2
MiniMax's M3 clears GPT-5.5 and Gemini 3.1 Pro on SWE-Bench Pro, hits 70.06% on OSWorld for computer use, and delivers 15x decoding speedup at 1M tokens via a new sparse attention architecture. Open weights within 10 days.
Florida Files the First State Lawsuit Against OpenAI and Sam Altman — FSU Shooting Chat Logs Cited
Florida Attorney General James Uthmeier sued OpenAI and CEO Sam Altman on June 1 under the state's Deceptive and Unfair Trade Practices Act, alleging ChatGPT aided a mass shooter, drove users to suicide, and addicted minors. It is the first state-level lawsuit against OpenAI in the US.
Groq Raises $650M to Rebuild as Inference Neocloud After Nvidia Took Its IP and Founders
Six months after a $20B Nvidia licensing deal stripped Groq of its architecture, founders, and engineering leadership, the company is raising $650M from existing investors to build an AI inference cloud. The window to survive is 18 to 24 months.