Anthropic's Evan Hubinger: AI extinction risk >10% this decade, alignment unsolved
SatelliteNetSec · x · 2026-09-10
Responding to a debate over data-center construction, Anthropic researcher Evan Hubinger said the company sincerely believes AI could kill all humans — his personal estimate is >10% within the next decade. He says Anthropic is trying its best, but there is no plan yet to solve alignment for superintelligence and the field isn't clearly on track. The quote drew both sarcasm (the best case is mass job loss and oligarchy) and calls from open-source advocates for 100x more research through open models and training data.
More from AGI Musings
- After Planning Two Books with ChatGPT, He No Longer Feels the Need to Write Them — tinyfool · 2026-09-11
- Indie hackers aren't just engineers or marketers — AI lets one builder run the whole loop — alexmacgregor__ · 2026-09-11
- AI Safety Practitioner: Cheap Extinction Talk Has Turned the Public Anti-AI — Dr_Atoosa · 2026-09-11
- After 9 Years, Bots Finally Show Up on a Blogger's Bounty Page with AI Text — gleech · 2026-09-11
- Mozilla CTO calls for major pause on generative AI in schools, warns of losing a generation — Dan_Jeffries1 · 2026-09-11
- Hypothesis: ASI Has a Mathematical Incentive to Preserve Human Diversity — No_Cause_2731 · 2026-09-11