Anthropic alignment lead agrees AI "kill all humans" odds exceed 10% in next decade
firstadopter · x · 2026-09-14
- Former Anthropic employee Jacob Coxon claimed the probability that AI will "kill all humans" is greater than 10% in the next decade, and Anthropic's Alignment Science lead Evan Hubinger agreed.
- firstadopter blasts the claim as shameful and irresponsible: it amounts to saying any newborn faces a >10% chance of perishing at AI's hands before middle school.
- Such doomsday probability claims are leaping into the mainstream, reigniting debate over the bounds of AI-safety rhetoric.
Related event: Anthropic Alignment Lead Backs 10% Extinction Risk Claim(2 posts)→
More from AGI Musings
- AI alignment must account for changing and contested values, argues Dylan Hadfield-Menell — dhadfieldmenell · 2026-09-14
- Boltz-2 run 100 million times: Recursion researcher builds a minimal virtual cell — HannesStaerk · 2026-09-14
- Naval amplifies AGI economics paper: verification, not intelligence, is the binding constraint — naval · 2026-09-14
- What If AI Was Owned Collectively? The Building Blocks of Decentralized AI — Admirable_Wasabi_732 · 2026-09-14
- Reasoning models' success makes denying computationalism ever harder — Aaroth · 2026-09-14
- Aaron Roth: worst-case complexity is a bad argument against building real AI — Aaroth · 2026-09-14