Anthropic Researcher Puts AI Extinction Risk Above 10% This Decade, Admits No Alignment Plan
SIGKITTEN · x · 2026-09-09
Anthropic researcher EvanHub publicly stated that the team genuinely believes AI could kill all humans, personally estimating the risk at over 10% within the next decade. He says Anthropic is trying its best, but has no plan yet to solve alignment for superintelligence and is not clearly on track. Critics replied that the statement amounts to admitting failure at the job while making the employer look bad, with zero actionable outcome.
More from AGI Musings
- Sutskever: Human Collaboration Is a Superintelligence Technology; Verdon Retorts With Agents — beffjezos · 2026-09-09
- Ex-OpenAI VP Miles Brundage: if you're considering leaving the AI industry, you probably should — Miles_Brundage · 2026-09-09
- Mathematician on OpenAI's Navier-Stokes push: picking the right question beats solving it — PTenigma · 2026-09-09
- Senate Tech Committee Schedules Zero AI Hearings for Next 4 Months — austinc3301 · 2026-09-09
- 'Attach sensors, train RL on the real world — then we're done': a one-line AGI thesis — Paimaamu · 2026-09-09
- AI 2027 was mocked as too fast — GPT-6 Astra just landed on its curve — Confident_Salt_8108 · 2026-09-09