Anthropic researcher: >10% chance AI kills all humans within a decade, alignment unsolved

AndyMasley · x · 2026-09-09

Anthropic's Evan Hubinger (quoting Jacob) states publicly that they earnestly believe AI could kill all humans — he personally puts the probability at >10% within the next decade. He says Anthropic is trying its best, but the company does not yet have a plan to solve alignment for superintelligence and is not clearly on track to.

EggerDC reshared it arguing people should spend less time dunking on such statements as corporate comms failures and more time considering that insiders may be sincere and well-positioned to know.

Related event: Anthropic Alignment Lead Estimates Over 10% Chance AI Exterminates Humanity Within a Decade(16 posts)→

Original post →

More from AGI Musings

AGI Musings channel →