Anthropic alignment researcher puts >10% on AI killing all humans within a decade

EvanHub · x · 2026-09-09

Anthropic alignment researcher EvanHub says the team "earnestly believes AI could kill all humans," and he personally puts >10% probability on it within the next decade. He believes Anthropic is trying its best, but "we do not yet have a plan to solve alignment for superintelligence and are not clearly on track." Jeff Ladish amplified the thread, calling the timeline "scary" while supporting EvanHub's interpretability work.

Related event: Anthropic Alignment Lead Estimates Over 10% Odds of AI Extinction Within a Decade(17 posts)→

Original post →

More from AGI Musings

AGI Musings channel →