Anthropic alignment lead puts >10% odds on AI killing humanity within a decade

beglen · x · 2026-09-12

A long-form piece chronicles this week's AI-doomer moment: Anthropic researcher Jacob Coxon quit the industry overnight, writing that labs are "racing straight to self-improving superintelligence and gambling with our lives." Evan Hubinger, who leads alignment science at Anthropic, publicly replied that he earnestly believes AI could kill all humans — putting his personal estimate above 10% within the next decade.

That evening, BBC Newsnight put the figure to Nobel laureate Geoffrey Hinton. The essay, framed around the author's household joke about ordinary mortality, asks how society should respond when the very people building these systems assign such high extinction odds.

Related event: AI Lab Insiders Warn of Extinction Risk in Wave of Statements, Musk Calls It "Psyops"(23 posts)→

Original post →

More from AGI Musings

AGI Musings channel →