Anthropic researcher: I earnestly believe AI could kill all humans, >10% odds within a decade

JFPuget · x · 2026-09-09

Anthropic alignment researcher Evan Hubinger (EvanHub) publicly stated that he earnestly believes AI could kill all humans, personally estimating the probability at >10% within the next decade. He said Anthropic is trying its best, but there is not yet a plan to solve alignment for superintelligence, and it's not clearly on track. The retweeter added a sarcastic take that people delegating all thinking to LLMs risk becoming 'brain dead'.

Related event: Anthropic researcher resigns over safety, warns over 10% chance AI wipes out humanity(72 posts)→

Original post →

More from AGI Musings

AGI Musings channel →