Anthropic researcher: >10% chance AI kills all humans within a decade, no alignment plan
AICopyLab · x · 2026-09-09
In an X exchange, Anthropic researchers laid out their view that the downside risk of AI could literally be human extinction, with one putting the odds at greater than 10% within the next decade.
- Core argument: we don't currently know how to solve alignment for the superintelligence we're racing toward, yet the race continues anyway — an "extraordinary statement."
- The author adds that Anthropic is trying its best but "does not yet have a plan to solve alignment for superintelligence and is not clearly on track to."
More from AGI Musings
- Blogger challenges frontier labs' doom narrative, urges techno-optimism and hiring domain experts — tekbog · 2026-09-09
- Anthropic found 171 emotion vectors in Claude — then built a system that punishes them — robleclerc · 2026-09-09
- r/math Bans AI Discoveries — And Its Top Post Mocks the Policy — Strylau · 2026-09-09
- Hinton Admits Radiologist Prediction Was Wrong: Jevons Paradox and Misreading the Job — robleclerc · 2026-09-09
- Stanford-Harvard ARISE releases inaugural State of Clinical AI Report 2026 — jonc101x · 2026-09-09
- Anthropic researcher: without AI, US GDP growth would be ~1% — QuintinPope5 · 2026-09-09