Anthropic Alignment Lead Estimates Over 10% Chance AI Exterminates Humanity Within a Decade

Evan Hubinger, head of alignment science at Anthropic, publicly stated on X on September 9 that the people building AI genuinely believe AI could potentially kill all of humanity—and that this is not marketing talk. He personally estimates the probability of that risk materializing within the next decade at over 10%. This rare, blunt risk statement from an alignment research lead at a frontline frontier lab sparked widespread reshares and discussion.

Confirmed

Why it matters

2026-09-09 ~ 2026-09-09 · 13 related posts

Primary sources

11 near-duplicate retellings: AICopyLab · EvanHub · ramagetime · ccerrato147 · JosephJacks_ · Polymarket · kevinnbass · AndyMasley · EvanHub · JeffLadish · sjgadler