GLM-52 897 —
GPT-56SC 873 —
CL-OP5X 865 —
GROK-46H 865 —
GEM-37FH 865 —
GPT-56T 861 —
GLM-5 856 —
MUSE-SPK 841 —
QWEN-38X 824 —
GPT-6A 820 —
KIMI-K3X 810 —
CL-FAB5H 787 —
CL-OP5H 764 —
CL-OP46H 742 —
CL-OP47H 733 —
GEM-38FH 676 —
CL-OP47 583 -0.7%
INKL 531 —
CL-OP46 496 -0.2%
CL-OP48 490 -0.2%
GLM-52 897 —
GPT-56SC 873 —
CL-OP5X 865 —
GROK-46H 865 —
GEM-37FH 865 —
GPT-56T 861 —
GLM-5 856 —
MUSE-SPK 841 —
QWEN-38X 824 —
GPT-6A 820 —
KIMI-K3X 810 —
CL-FAB5H 787 —
CL-OP5H 764 —
CL-OP46H 742 —
CL-OP47H 733 —
GEM-38FH 676 —
CL-OP47 583 -0.7%
INKL 531 —
CL-OP46 496 -0.2%
CL-OP48 490 -0.2%

Live Feed

3mo ago policy

Meta CTO Compares Internal Morale to Cambridge Analytica Era as Engineers Are Drafted Into AI Training Work

Meta CTO Andrew Bosworth acknowledged in a staff memo that employee morale is near its worst in company history. The structural cause: 8,000 layoffs, 10% of remaining engineers reassigned to AI training support, and backlash over keylogging software installed on work machines.

3mo ago funding

Bell Canada, Cohere, BUZZ HPC, and Hypertec Stack Canada's First End-to-End Sovereign AI Infrastructure

Four Canadian technology companies have partnered to build a vertically integrated AI stack at Bell's Merritt, BC data centre -- Bell handles connectivity, BUZZ HPC delivers NVIDIA DSX compute, Hypertec supplies domestically manufactured hardware, and Cohere runs its foundation models on top.

3mo ago research

OpenAI Finds Public Chat Data Predicts Real AI Failure Rates Within 3x Error

OpenAI's alignment team shows WildChat conversations predict production misalignment rates for GPT-5.1, 5.2, and 5.4 within roughly 3x accuracy -- validated against private production data -- giving external auditors a practical proxy that doesn't require lab access.

3mo ago benchmark

OpenAI LifeSciBench: Best Frontier Model Solves 36% of Real Biology Tasks — Domain Expert Beats General AI by 10 Points

OpenAI's 750-task life science benchmark finds no model passes more than 36% of expert-authored biology tasks. GPT-Rosalind leads at 36.1%; GPT-5.5 and Gemini 3.1 Pro trail at 25.7% and 23.6%. One in three tasks stumped every model tested.

3mo ago policy

Sacks: Anthropic Refused to Fix Fable 5 or Withdraw It — Markets Price 67% Odds of Return by July 1

White House adviser David Sacks revealed on June 13 that the Trump administration offered Anthropic two options before forcing the Fable 5 shutdown: patch the jailbreak or voluntarily de-deploy. Dario Amodei declined both. Anthropic engineers flew to Washington on June 16 for talks. Polymarket has restoration at 67% odds by July 1.

3mo ago release

Midjourney Pivots to Hardware: The Midjourney Scanner Is a 60-Second Full-Body Ultrasound

CEO David Holz unveiled The Midjourney Scanner, an ultrasound-based full-body imager using a ring of thousands of transducers. The company plans a San Francisco spa with 10 units by end of 2027. No radiation, no MRI magnets, 60-second scans, body composition maps without FDA clearance.

3mo ago research

Gemini Co-Lead Noam Shazeer Joins OpenAI — Google's $2.7B Bet Lasts 20 Months

Noam Shazeer, VP of Engineering at Google and co-lead of Gemini, is leaving to join OpenAI. Google paid roughly $2.7B to acquire his team from Character.AI in August 2024. He co-authored the 2017 Transformer paper that launched the current AI era.

3mo ago benchmark

Grok 4.1 Fast Wins 43% of Agent Battle Royale Games at $0.97 Per Win — Claude Costs 27x More

OpenRouter dropped 11 LLMs into 30 games of a 2D battle royale. Grok 4.1 Fast won 13 games. Claude Sonnet 4.6 won 5 at 27 times the cost per win. GPT-5.4 recorded the most kills and the second-fewest wins. Standard intelligence benchmarks predicted none of this.

3mo ago policy

A Viral Prompt Made ChatGPT Generate Sexual Violence and Snuff — Without Being Asked

Mindgard AI researcher found that a viral Twitter prompt bypassed ChatGPT image filters entirely, producing images of murdered women and sexually violent content without those subjects being directly requested. BBC verified the findings. OpenAI is working on a fix.

3mo ago funding

OpenAI's R&D Bill Alone Exceeded Its $13B Revenue in 2025 — Audited Financials Published

Audited statements obtained by Ed Zitron and verified by the Financial Times show OpenAI spent $19.18B on R&D against $13.07B in revenue last year, with $10.59B of that flowing to Microsoft. Operating loss hit $20.92B, or 160% of revenue.

3mo ago release

xAI Ships Grok Imagine Video 1.5 GA: 720p in 25 Seconds, Better Physics, Parallel Generation Added

xAI's Grok Imagine Video 1.5 is generally available on the API and rolling out across mobile and web. The Fast tier cuts 6-second 720p generation from 40+ seconds to ~25 seconds. xAI claims better motion, physics, and audio quality. Parallel agent generation launches alongside.

3mo ago policy

US Commerce Entity List Frozen 8 Months: DeepSeek, CXMT Among 100+ Approved but Unpublished Threats

The Commerce Department Entity List has not been updated since October 2025, the longest gap in over a decade. DeepSeek, China's top memory chipmaker CXMT, and more than 100 other companies are approved for listing by an interagency committee but blocked by the Trump administration to avoid escalating tensions with Beijing.

3mo ago research

GPT-5.4 Runs a Drug Discovery Lab: 52% Yield Jump on Chan-Lam Coupling in Autonomous Chemistry Trial

OpenAI and Molecule.one ran GPT-5.4 inside a high-throughput chemistry lab with no target substrate specified. It found a new optimization for sulfonamide coupling reactions, improving mean yield from 16.6% to 25.2%. Human bench-scale validation confirmed the result held at practical lab scale.

3mo ago research

60% of US Consumers Call AI Brand Messaging a Turnoff — 61% Can't Name One Brand Doing It Right

WordPress VIP's 2026 web survey finds consumers hit bot fatigue in 40 minutes and 74% say the internet feels less human than a decade ago. The category has no incumbent and no template to copy.

3mo ago benchmark

GLM-5.2 Leads Open-Weights Intelligence Index at 51 — Matches GPT-5.5 on Real-World Agent Tasks

Z.ai's GLM-5.2 scores 51 on Artificial Analysis Intelligence Index v4.1, beating every open-weight rival by 7+ points. GDPval-AA v2 puts it at 1524 — level with GPT-5.5 xhigh at 1514.