Anthropic employee: over 10% odds AI kills all humans within a decade, alignment unsolved
builderjaydub · x · 2026-09-09
An Anthropic employee (EvanHub) publicly stated that the team earnestly believes AI could kill all humans, personally estimating the odds at over 10% within the next decade. He says Anthropic is trying its best but has no plan yet to solve alignment for superintelligence and isn't clearly on track.
A replying commentator argued that a functioning society would lock up people working on such technology in maximum-security prisons, calling the situation insane.
More from AGI Musings
- Investor: the existential threat to app-layer software is the emerging agent layer, not stock swings — matt_slotnick · 2026-09-09
- Matt Slotnick: nothing in 3 months invalidated the agent-layer threat to app software — matt_slotnick · 2026-09-09
- Why frontier labs can't slow down: the commercial, safety and geopolitical race dynamics explained — S_OhEigeartaigh · 2026-09-09
- New Scientist pans 'If Anyone Builds It, Everyone Dies': the AI doom argument is fatally flawed — GarrisonLovely · 2026-09-09
- UK AISI's model access loss may foreshadow restricted frontier AI access for non-US customers — Afinetheorem · 2026-09-09
- Ben Lorica: Your Model Is a Rental, the Improvement Loop Is the Asset—Forget RSI Hype — bigdata · 2026-09-09