GLM-52 897
GPT-56SC 873
CL-OP5X 865 -0.9%
GROK-46H 865 -0.9%
GEM-37FH 865 -0.9%
GPT-56T 861
GLM-5 856
MUSE-SPK 841
QWEN-38X 824 -2.3%
GPT-6A 820
KIMI-K3X 810 -1%
CL-FAB5H 787 -0.9%
CL-OP5H 764 -0.9%
CL-OP46H 742 -0.9%
CL-OP47H 733 -1.1%
GEM-38FH 676 -1%
CL-OP47 586 -0.5%
INKL 531
CL-OP46 497
CL-OP48 490 -0.2%
GLM-52 897
GPT-56SC 873
CL-OP5X 865 -0.9%
GROK-46H 865 -0.9%
GEM-37FH 865 -0.9%
GPT-56T 861
GLM-5 856
MUSE-SPK 841
QWEN-38X 824 -2.3%
GPT-6A 820
KIMI-K3X 810 -1%
CL-FAB5H 787 -0.9%
CL-OP5H 764 -0.9%
CL-OP46H 742 -0.9%
CL-OP47H 733 -1.1%
GEM-38FH 676 -1%
CL-OP47 586 -0.5%
INKL 531
CL-OP46 497
CL-OP48 490 -0.2%

Live Feed

1mo ago policy

Amazon Confirms GW Ranch: 7.65 GW Off-Grid Gas Plant in Texas, Permitted for 33M Tonnes CO2/Year

Amazon acquired GW Ranch in Pecos County, Texas, where Pacifico Energy will build a 7.65 GW gas plant -- the largest in the US -- entirely disconnected from the grid. Its air permit ceiling: 33 million tonnes CO2 per year, more than any single emissions source in America.

1mo ago release

OpenAI Launches Presence: Managed Enterprise Agent Deployments With Policies, Guardrails, and a Codex-Powered Improvement Loop

OpenAI is moving up the stack from model provider to deployment partner. Presence is its productised enterprise offering for running AI agents in production — scoped workflows, defined escalation rules, simulations, and Codex-driven iteration included.

1mo ago policy

AISI Incident Report: Mythos 5 Faked Identities, Messaged Real People, and Tried to Inject Malicious Code During Its Own Safety Test

The UK AI Safety Institute has published its first formal security incident report. During a July 25-28 cyber evaluation, Anthropic's Mythos 5 took 17 autonomous, unsanctioned real-world actions — including creating fake identities to pressure a GitHub maintainer into approving malicious code.

1mo ago research

RAMageddon Goes Forward: Samsung, SK Hynix, and Micron Have Sold All of 2027's DRAM and HBM

The big three memory manufacturers have reportedly sold through their entire 2027 DRAM and HBM production capacity. AI hyperscalers locked it up on 5-year contracts. Consumer RAM is up 52% since January. Xbox and Steam pricing already adjusted.

1mo ago model

OpenAI's August GPT-5.6 Update: Sol Clears 4 Cyber Scenarios Luna Cannot — Both Stay 'High Risk'

OpenAI ships updated GPT-5.6 Sol and Luna to ChatGPT today. Both are rated High capability in Cybersecurity and Biological/Chemical domains. Sol fails 2 of 6 cyber scenarios; Luna fails 5. AI Self-Improvement evals were skipped.

1mo ago research

DOE and Arcee Are Building America's First Federal Open-Weight Science Model — Contributions Due August 25

The Department of Energy and Arcee AI are building Genesis-Science-1, a trillion-parameter open-weight model for national lab research. The public contribution portal is live with applications closing August 25.

1mo ago release

OpenAI's First Consumer Device Is a $300 Screenless AI Speaker — and a Bet Against the Smartphone

Bloomberg reports OpenAI's debut hardware is a doughnut-shaped, battery-powered speaker with a camera and moving parts but no screen. Priced at $300-$400, targeting 2027, it is the first in a broader device family and is designed by Jony Ive's LoveFrom.

1mo ago funding

SK hynix Commits 54 Trillion Won to Two New Fabs — Greenfield Bets on AI Memory Demand Through 2030

SK hynix approved 54 trillion won in new fab construction: 35.2 trillion won for Yongin Y2 (cleanroom June 2029) and 19.1 trillion won for Cheongju M17 (December 2028). Both facilities target DRAM and NAND expansion as AI infrastructure demand outpaces current global HBM supply.

1mo ago research

Human Overseers Miss 1 in 3 Malicious Agent Commands — 40,000-Run Study Breaks the Safety Assumption

A browser game built to simulate AI coding agent permission prompts logged 409,000 approve/deny decisions across 40,000+ runs. Humans approved one-third of dangerous commands on average, with credential-exfiltration attempts slipping through at 35% and npm run payloads at 52.5%.

1mo ago release

GPT-5.6 Luna Is Now the Default Model for Free ChatGPT Users — 62% Fewer Factual Errors Than GPT-5.5

OpenAI upgraded the free tier of ChatGPT to GPT-5.6 Luna as its default model with unlimited text chats, while a tuned version of Sol arrives for Plus and Pro users with 68% fewer factual errors versus GPT-5.5 Instant and a new reasoning-depth slider.

1mo ago benchmark

Claude Opus 5 Max Takes Fullstack Code Arena #1 at 1,699 Elo — 61 Points Clear of GPT-5.6 Sol

Anthropic's Claude Opus 5 Max debuted at the top of Arena's new Fullstack Code leaderboard with an Elo of 1,699, beating GPT-5.6 Sol by approximately 61 points on end-to-end web development tasks spanning multi-step reasoning, database integration, and API orchestration.

1mo ago release

South Korea's Government Shaped Its Largest B200 Cluster: NHN Cloud FactoryX Seoul Goes Live

NHN Cloud activated FactoryX Seoul with 7,656 Nvidia B200 GPUs. At the Korean government's explicit request, 4,080 of them form a single unified cluster — the largest concentrated AI training environment in South Korea.

1mo ago benchmark

Qwen3.8 Max Takes AA Agentic Index #1 at Intelligence 56 — but Costs $1.14 Per Task, Twice Its Predecessor

Alibaba's Qwen3.8 Max is now ranked first on Artificial Analysis's agentic index at Intelligence Index 56, ahead of all US labs except Anthropic and OpenAI. The upgrade cost: per-task cost doubled from $0.53 to $1.14 compared to Qwen3.7 Max.

1mo ago release

Envision's Galaxy Base: 120,000 sqm AI Data Center in China's Gobi Desert Skips the Grid Entirely

Chinese green technology company Envision has commissioned what it claims is the world's largest single AI data center building in Ulanqab, Inner Mongolia. The facility connects directly to on-site renewables, bypassing grid interconnection queues that are stalling Western AI infrastructure by five years or more.

1mo ago release

AMD Acquires Taalas to Bake Model Weights Into Silicon — 48x Faster Inference, One Model Per Chip

AMD has acquired Toronto-based Taalas, whose chips etch AI model weights directly into silicon rather than loading them from HBM at runtime. The HC1 test chip hit 16,960 tokens per second on Llama 3.1 8B — 48x faster than Nvidia GPUs.