Anthropic researcher Evan Hubinger puts >10% odds on AI killing all humans within a decade
JoshPurtell · x · 2026-09-09
Anthropic alignment researcher Evan Hubinger said on X that the team sincerely believes AI could kill all humans, putting his personal estimate at >10% within the next decade. He admitted Anthropic is trying its best but has no plan to solve alignment for superintelligence and isn't clearly on track to.
The statement drew sharp backlash: user JoshPurtell blasted him for claiming doom is imminent while making "absolutely fuck all progress" on alignment, telling him to resign. The exchange highlights the ongoing tension inside the safety community between doom warnings and slow research progress.
More from AGI Musings
- Math Bodies Must Decide What Counts as Proof as AI Slop Proofs Flood In — thebasepoint · 2026-09-09
- Ray Dalio: every tech boom creates a bubble — the miracle vs. the investment — RayDalio · 2026-09-09
- AI Video Now Generates Faster Than You Can Watch, Endless Slop Incoming — michalmalewicz · 2026-09-09
- Why AI lab researchers who believe in ~10% extinction risk keep working there — sebkrier · 2026-09-09
- Dwarkesh: start your AI-era institution now — it could become society's default — luke_drago_ · 2026-09-09
- Blogger predicts 10,000 agents will crack Millennium math problem by late 2026, 1M agents to attack cancer in 2027 — Dr_Singularity · 2026-09-09