Anthropic alignment lead Evan Hubinger: >10% chance AI kills all humans within a decade
No-Meringue5867 · reddit · 2026-09-09
Evan Hubinger, Alignment Science lead at Anthropic, publicly stated that the team earnestly believes AI could kill all humans — he personally puts the risk at over 10% within the next decade — and admitted there is not yet a plan to solve alignment for superintelligence, nor is the field clearly on track to one.
More from AGI Musings
- Blogger proposes study measuring how often AI trend-mockers get proven wrong — moultano · 2026-09-09
- Why AI Hasn't Boosted Growth Yet: It's a Function of Global Inference Capacity — zephyr_z9 · 2026-09-09
- Math lacks metascience tradition, making AI slop claims weakly grounded, researcher argues — RexDouglass · 2026-09-09
- AI practitioner mocks AGI hype: researchers can't even deploy AI securely in labs — iamKierraD · 2026-09-09
- OpenAI Solved a Millennium Prize Problem — So Why Is Software Still Buggy? — ziv_ravid · 2026-09-09
- From 'crackpot' to plausible: AI solving Navier-Stokes flipped in months — EigenGender · 2026-09-09