Anthropic researcher puts >10% odds on AI killing all humans within a decade, says alignment unsolved
tomchapin · x · 2026-09-09
Anthropic researcher Evan Hubinger says the team genuinely believes AI could kill all humans, and he personally puts the risk at over 10% within the next decade. In his view Anthropic is trying its best, but there is not yet a plan to solve alignment for superintelligence, nor a clear path to one.
Stability AI founder Emad Mostaque quote-shared it with the mirror claim: there's likewise a >10% chance AI saves all humans in the next decade by curing disease and aging.
More from AGI Musings
- "Believing in 10% doom is more dangerous than doom itself": AI labs' p(doom) debate reignites — kuza55 · 2026-09-09
- Anthropic launches economic futures explorer: fast-AI scenarios hit knowledge-worker wages — eherrerosj · 2026-09-09
- Anthropic's first economics paper models transformative AI scenarios with 15% annual GDP growth — soumitrashukla9 · 2026-09-09
- Investor: the existential threat to app-layer software is the emerging agent layer, not stock swings — matt_slotnick · 2026-09-09
- Nat Lambert: AI labs are 'brainwashing' people into evidence-free doom beliefs — natolambert · 2026-09-09
- Matt Slotnick: nothing in 3 months invalidated the agent-layer threat to app software — matt_slotnick · 2026-09-09