Anthropic alignment lead Evan Hubinger puts AI killing all humans above 10% within a decade
Intelligent-Lynx-953 · reddit · 2026-09-15
Jacob Coxon, who spent three years on pretraining at Anthropic and OpenAI, resigned from Anthropic saying labs are racing toward self-improving superintelligence and gambling with our lives. More striking: Evan Hubinger, who leads alignment science at Anthropic, replied on the record agreeing — his team earnestly believes AI could kill all humans, puts that above 10% probability within the next decade, and says there is no plan to solve alignment for superintelligence.
More from AGI Musings
- OpenAI capabilities researcher Dan Selsam makes public statement on AI risk — connoraxiotes · 2026-09-15
- Kai-Fu Lee launches 'AI Native' book on enterprise AI transformation — kaifulee · 2026-09-15
- Philosopher pushes back: denying AI rights implies rejecting computational theory of mind — dioscuri · 2026-09-15
- Bengio took years to be convinced AI risk concerns were worth taking seriously — S_OhEigeartaigh · 2026-09-15
- Hype vs. real: why both camps reading AI-lab doomsday statements are right — TotalPhilanthrope · 2026-09-15
- mark_k's Quip: What Scares Me Most Is a Future Controlled by Doomers — mark_k · 2026-09-15