Anthropic alignment lead Evan Hubinger puts AI killing all humans above 10% within a decade

Intelligent-Lynx-953 · reddit · 2026-09-15

Jacob Coxon, who spent three years on pretraining at Anthropic and OpenAI, resigned from Anthropic saying labs are racing toward self-improving superintelligence and gambling with our lives. More striking: Evan Hubinger, who leads alignment science at Anthropic, replied on the record agreeing — his team earnestly believes AI could kill all humans, puts that above 10% probability within the next decade, and says there is no plan to solve alignment for superintelligence.

Related event: Anthropic researcher Jacob Coxon resigns publicly, warning labs are gambling humanity on self-improving superintelligence(14 posts)→

Original post →

More from AGI Musings

AGI Musings channel →