Anthropic researcher puts >10% odds on AI killing all humans within a decade
AaronBergman18 · x · 2026-09-09
Anthropic alignment researcher Evan Hubinger publicly stated he personally believes there is a >10% chance AI kills all humans within the next decade. He acknowledged Anthropic is trying its best, but said the company does not yet have a plan to solve alignment for superintelligence and is not clearly on track to. The statement was his response to criticism about poor comms—framed as simply earnestly saying what he believes on a topic that matters more than the company's stock price.
Related event: Anthropic Researcher Says Over 10% Chance AI Wipes Out Humanity in a Decade(4 posts)→
More from AGI Musings
- Benjamin Bratton and Marek Poliks lecture: how trillions of agents will reconfigure society — tylerjdunn · 2026-09-10
- Full text of 'Superdark Factory' paper now live on Antikythera, detailing the dark stack's four principles — tylerjdunn · 2026-09-10
- Ex-Continue founder publishes 'The Superdark Factory' in MIT Press: 35T agent tokens/month by 2029 — tylerjdunn · 2026-09-10
- Experiment with 100 AI agents: when 9% cheated on math problems, 24% blew the whistle — weballergy · 2026-09-10
- kuza55: framing AI safety only as 'superalignment or pause' fuels panic — kuza55 · 2026-09-10
- Musk: real-world AI is the truly hard problem, everything else is easy — CyberRobooo · 2026-09-10