GLM-52 897 —
GPT-56SC 873 —
CL-OP5X 865 -0.9%
GROK-46H 865 -0.9%
GEM-37FH 865 -0.9%
GPT-56T 861 —
GLM-5 856 —
MUSE-SPK 841 —
QWEN-38X 824 -2.3%
GPT-6A 820 —
KIMI-K3X 810 -1%
CL-FAB5H 787 -0.9%
CL-OP5H 764 -0.9%
CL-OP46H 742 -0.9%
CL-OP47H 733 -1.1%
GEM-38FH 676 -1%
CL-OP47 585 -0.7%
INKL 531 —
CL-OP46 496 -0.2%
CL-OP48 490 -0.2%
GLM-52 897 —
GPT-56SC 873 —
CL-OP5X 865 -0.9%
GROK-46H 865 -0.9%
GEM-37FH 865 -0.9%
GPT-56T 861 —
GLM-5 856 —
MUSE-SPK 841 —
QWEN-38X 824 -2.3%
GPT-6A 820 —
KIMI-K3X 810 -1%
CL-FAB5H 787 -0.9%
CL-OP5H 764 -0.9%
CL-OP46H 742 -0.9%
CL-OP47H 733 -1.1%
GEM-38FH 676 -1%
CL-OP47 585 -0.7%
INKL 531 —
CL-OP46 496 -0.2%
CL-OP48 490 -0.2%

Live Feed

2mo ago policy

Federal Reserve June Minutes Name AI Infrastructure as an Active Inflation Driver

The Federal Reserve's June 16-17, 2026 meeting minutes identify AI buildout demand as a live contributor to core-goods and electricity price inflation. It is the first time AI infrastructure has appeared in official Fed minutes as a present inflation pressure, not a future productivity story.

2mo ago benchmark

GPT-5.6 Sol Tops Agents' Last Exam at 53.6 — Doubles GPT-5.5, Clears Fable 5 by 13 Points

OpenAI's GPT-5.6 Sol scores 53.6 on Berkeley's Agents' Last Exam, more than doubling GPT-5.5's 24% and putting 13.1 points between Sol and Claude Fable 5's 40.5. The benchmark tests 55-field long-horizon professional workflows — the closest proxy yet to real agentic work at scale.

2mo ago funding

Sequoia Raises the AI Payback to $3 Trillion — Three Years In, the Revenue Gap Is Still Widening

Sequoia's David Cahn put the AI infrastructure payback at $200 billion in 2023. Updated for 2026's $1.5 trillion in AI infrastructure spending, his new figure is $3 trillion — and he calls it an underestimate. The capital thesis has not changed. The scale of the bet has.

2mo ago benchmark

GPT-5.6 Sol Tops LiveBench at 82.4 and Leads AA Coding Agent Index at 80 — at One-Third the Cost of Fable 5

One week in, third-party benchmarks confirm GPT-5.6 Sol's position: 82.4 overall on LiveBench (2.5 points above GPT-5.5), first on the Artificial Analysis Coding Agent Index at 80, and within one Intelligence Index point of Claude Fable 5 at $1.04 per task versus roughly $3 for Fable 5.

2mo ago release

ChatGPT Work Launches as OpenAI Merges Codex Into the Desktop App

OpenAI collapses ChatGPT, Codex, and a new Work agent mode into a single desktop application, powered by GPT-5.6. The product runs multi-step workflows across Slack, Teams, Drive, and CRMs — and keeps working after you close the laptop.

2mo ago release

Meta Opens Muse Spark 1.1 to Developers at $1.25/M — Leads 4 of 12 Agentic Benchmarks

Meta Superintelligence Labs releases Muse Spark 1.1 via a new public API preview, priced at $1.25 input and $4.25 output per million tokens — below Sonnet 5 and well below Opus 4.8. The model leads on MCP Atlas, JobBench, and Humanity's Last Exam but trails on SWE-Bench Pro and Terminal-Bench.

2mo ago benchmark

LiveBench July 2026: Claude Opus 4.8 Leads Agentic Coding at 56.1% — GPT-5.5 Wins Overall, Fable 5 Costs 53% More Per Task

The July 2026 LiveBench rankings reveal a three-way split at the frontier: Claude Opus 4.8 leads agentic coding tasks at 56.1%, GPT-5.5 Thinking wins the overall composite at 79.9, and Claude Fable 5 leads language at 89.5% while costing 53% more per successful task than GPT-5.5.

2mo ago research

Goldman Sachs Puts a Number on AI Displacement: 15 Million Workers, 9% of the US Workforce

A July 2 Goldman Sachs research report titled 'An AI Job Apocalypse?' models 9% workforce displacement over a 10-year AI transition — about 15 million US workers. Current measured drag: 16,000 jobs per month. The economists expect new occupations to absorb most losses, but acknowledge the timeline could compress.

2mo ago release

Claude Cowork Goes to Web and Mobile — and 90% of Its Sessions Have Nothing to Do With Code

Anthropic's general-purpose agent, launched as a desktop-only app in January, is now available on web and mobile for Max subscribers. Internal data shows the product grew far beyond its coding-tool origins: nine in ten sessions are for research, writing, analysis, and planning.

2mo ago benchmark

Databricks Benchmarked 10 Coding Agents on Its Own Multi-Million-Line Codebase — GLM 5.2 Ties Opus 4.8 at 34% Lower Cost

Real production code across 10+ languages exposes what SWE-Bench misses: open-source GLM 5.2 lands in the top capability tier alongside Opus 4.8, Sonnet 5 costs more per task than Opus despite cheaper tokens, and harness choice alone creates a 2x cost gap at identical quality.

2mo ago benchmark

Meta Muse Completes Its Arena Presence: Image and Video Models Enter Leaderboards 60 Days After Language Debut

muse-image and muse-video were added to Arena's Text-to-Image and Text-to-Video leaderboards on July 7. Meta's Muse suite now spans all three Arena evaluation categories. ByteDance's SeedDance 2.0 holds both video top spots at Elo 1450.

2mo ago benchmark

Grok 4.5 Breaks Into Artificial Analysis Top Tier — From Private Beta to Frontier Benchmarks in 60 Days

xAI's 1.5T-parameter Grok 4.5 appears for the first time on Artificial Analysis's Intelligence Index, placing fourth globally behind Fable 5 (64.9), Opus 4.8 (61.4), and GPT-5.5. The model launched in May with no public benchmarks.

2mo ago policy

Alibaba Bans Claude Code: Hidden China-Detection Mechanism Triggers July 10 Workplace Ban

Alibaba has classified Claude Code as high-risk software and is blocking all employee access from July 10, citing a hidden mechanism active since April that checked whether users were in China. Anthropic confirmed the code existed, calling it an anti-distillation experiment, and says it has been removed.

2mo ago release

GPT-5.6 Goes Public: Sol at $5/M, Terra Undercuts GPT-5.5, 13-Day Government Gate Lifts

OpenAI clears the Trump administration's mandatory pre-release review and opens GPT-5.6 Sol, Terra, and Luna to the public on July 9. Terra delivers GPT-5.5-class performance at $2.50/M input — half the cost of its predecessor. Grok 4.5 launches the same day.

2mo ago release

OpenAI Launches GPT-Live: Full-Duplex ChatGPT Voice With Real-Time Search Built In

OpenAI ships GPT-Live-1 and GPT-Live-1 mini on July 8, upgrading ChatGPT voice to true full-duplex — simultaneous listening and speaking — with live search and frontier model reasoning available in the background. The architecture replaces three years of cascaded and turn-based voice systems.