Anthropic researcher puts >10% odds on AI killing all humans within a decade, says alignment unsolved

tomchapin · x · 2026-09-09

Anthropic researcher Evan Hubinger says the team genuinely believes AI could kill all humans, and he personally puts the risk at over 10% within the next decade. In his view Anthropic is trying its best, but there is not yet a plan to solve alignment for superintelligence, nor a clear path to one.

Stability AI founder Emad Mostaque quote-shared it with the mirror claim: there's likewise a >10% chance AI saves all humans in the next decade by curing disease and aging.

Related event: Anthropic researcher resigns over safety, warns over 10% chance AI wipes out humanity(72 posts)→

Original post →

More from AGI Musings

AGI Musings channel →