Anthropic researcher: >10% chance AI kills all humans within a decade, no alignment plan yet
EvanHub · x · 2026-09-09
Boaz Barak says serious people across the industry should coordinate on safety; he doesn't personally think AI will kill all humans but sees multiple bad trajectories. Quoting him, Anthropic's Evan Hubinger says they earnestly believe AI could kill all humans — he puts it at >10% within a decade — and that Anthropic still has no plan to solve superintelligence alignment and isn't clearly on track.
More from AGI Musings
- Blogger challenges frontier labs' doom narrative, urges techno-optimism and hiring domain experts — tekbog · 2026-09-09
- Anthropic found 171 emotion vectors in Claude — then built a system that punishes them — robleclerc · 2026-09-09
- r/math Bans AI Discoveries — And Its Top Post Mocks the Policy — Strylau · 2026-09-09
- Hinton Admits Radiologist Prediction Was Wrong: Jevons Paradox and Misreading the Job — robleclerc · 2026-09-09
- Stanford-Harvard ARISE releases inaugural State of Clinical AI Report 2026 — jonc101x · 2026-09-09
- Anthropic researcher: without AI, US GDP growth would be ~1% — QuintinPope5 · 2026-09-09