Anthropic Alignment Lead: Over 10% Chance AI Kills All Humans Within a Decade

austinc3301 · x · 2026-09-09

Evan Hubinger, alignment lead at Anthropic, said the team genuinely believes AI could kill all humans — he personally puts the risk above 10% within the next decade. He acknowledged Anthropic is trying its best but admitted the company does not yet have a plan to solve alignment for superintelligence and is not clearly on track to.

The remarks were amplified by hecubiandevil with the note that this is Anthropic's alignment lead speaking.

Related event: Anthropic Alignment Lead Sees Over 10% Chance AI Destroys Humanity Within a Decade(30 posts)→

Original post →

More from AGI Musings

AGI Musings channel →