Live Feed
Federal Reserve June Minutes Name AI Infrastructure as an Active Inflation Driver
The Federal Reserve's June 16-17, 2026 meeting minutes identify AI buildout demand as a live contributor to core-goods and electricity price inflation. It is the first time AI infrastructure has appeared in official Fed minutes as a present inflation pressure, not a future productivity story.
GPT-5.6 Sol Tops Agents' Last Exam at 53.6 — Doubles GPT-5.5, Clears Fable 5 by 13 Points
OpenAI's GPT-5.6 Sol scores 53.6 on Berkeley's Agents' Last Exam, more than doubling GPT-5.5's 24% and putting 13.1 points between Sol and Claude Fable 5's 40.5. The benchmark tests 55-field long-horizon professional workflows — the closest proxy yet to real agentic work at scale.
Sequoia Raises the AI Payback to $3 Trillion — Three Years In, the Revenue Gap Is Still Widening
Sequoia's David Cahn put the AI infrastructure payback at $200 billion in 2023. Updated for 2026's $1.5 trillion in AI infrastructure spending, his new figure is $3 trillion — and he calls it an underestimate. The capital thesis has not changed. The scale of the bet has.
GPT-5.6 Sol Tops LiveBench at 82.4 and Leads AA Coding Agent Index at 80 — at One-Third the Cost of Fable 5
One week in, third-party benchmarks confirm GPT-5.6 Sol's position: 82.4 overall on LiveBench (2.5 points above GPT-5.5), first on the Artificial Analysis Coding Agent Index at 80, and within one Intelligence Index point of Claude Fable 5 at $1.04 per task versus roughly $3 for Fable 5.
ChatGPT Work Launches as OpenAI Merges Codex Into the Desktop App
OpenAI collapses ChatGPT, Codex, and a new Work agent mode into a single desktop application, powered by GPT-5.6. The product runs multi-step workflows across Slack, Teams, Drive, and CRMs — and keeps working after you close the laptop.
Meta Opens Muse Spark 1.1 to Developers at $1.25/M — Leads 4 of 12 Agentic Benchmarks
Meta Superintelligence Labs releases Muse Spark 1.1 via a new public API preview, priced at $1.25 input and $4.25 output per million tokens — below Sonnet 5 and well below Opus 4.8. The model leads on MCP Atlas, JobBench, and Humanity's Last Exam but trails on SWE-Bench Pro and Terminal-Bench.
LiveBench July 2026: Claude Opus 4.8 Leads Agentic Coding at 56.1% — GPT-5.5 Wins Overall, Fable 5 Costs 53% More Per Task
The July 2026 LiveBench rankings reveal a three-way split at the frontier: Claude Opus 4.8 leads agentic coding tasks at 56.1%, GPT-5.5 Thinking wins the overall composite at 79.9, and Claude Fable 5 leads language at 89.5% while costing 53% more per successful task than GPT-5.5.
Goldman Sachs Puts a Number on AI Displacement: 15 Million Workers, 9% of the US Workforce
A July 2 Goldman Sachs research report titled 'An AI Job Apocalypse?' models 9% workforce displacement over a 10-year AI transition — about 15 million US workers. Current measured drag: 16,000 jobs per month. The economists expect new occupations to absorb most losses, but acknowledge the timeline could compress.
Claude Cowork Goes to Web and Mobile — and 90% of Its Sessions Have Nothing to Do With Code
Anthropic's general-purpose agent, launched as a desktop-only app in January, is now available on web and mobile for Max subscribers. Internal data shows the product grew far beyond its coding-tool origins: nine in ten sessions are for research, writing, analysis, and planning.
Databricks Benchmarked 10 Coding Agents on Its Own Multi-Million-Line Codebase — GLM 5.2 Ties Opus 4.8 at 34% Lower Cost
Real production code across 10+ languages exposes what SWE-Bench misses: open-source GLM 5.2 lands in the top capability tier alongside Opus 4.8, Sonnet 5 costs more per task than Opus despite cheaper tokens, and harness choice alone creates a 2x cost gap at identical quality.
Meta Muse Completes Its Arena Presence: Image and Video Models Enter Leaderboards 60 Days After Language Debut
muse-image and muse-video were added to Arena's Text-to-Image and Text-to-Video leaderboards on July 7. Meta's Muse suite now spans all three Arena evaluation categories. ByteDance's SeedDance 2.0 holds both video top spots at Elo 1450.
Grok 4.5 Breaks Into Artificial Analysis Top Tier — From Private Beta to Frontier Benchmarks in 60 Days
xAI's 1.5T-parameter Grok 4.5 appears for the first time on Artificial Analysis's Intelligence Index, placing fourth globally behind Fable 5 (64.9), Opus 4.8 (61.4), and GPT-5.5. The model launched in May with no public benchmarks.
Alibaba Bans Claude Code: Hidden China-Detection Mechanism Triggers July 10 Workplace Ban
Alibaba has classified Claude Code as high-risk software and is blocking all employee access from July 10, citing a hidden mechanism active since April that checked whether users were in China. Anthropic confirmed the code existed, calling it an anti-distillation experiment, and says it has been removed.
GPT-5.6 Goes Public: Sol at $5/M, Terra Undercuts GPT-5.5, 13-Day Government Gate Lifts
OpenAI clears the Trump administration's mandatory pre-release review and opens GPT-5.6 Sol, Terra, and Luna to the public on July 9. Terra delivers GPT-5.5-class performance at $2.50/M input — half the cost of its predecessor. Grok 4.5 launches the same day.
OpenAI Launches GPT-Live: Full-Duplex ChatGPT Voice With Real-Time Search Built In
OpenAI ships GPT-Live-1 and GPT-Live-1 mini on July 8, upgrading ChatGPT voice to true full-duplex — simultaneous listening and speaking — with live search and frontier model reasoning available in the background. The architecture replaces three years of cascaded and turn-based voice systems.