Anthropic alignment researcher: >10% chance AI kills all humans within a decade
kevinnbass · x · 2026-09-09
Anthropic's Evan Hubinger: 'We really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade... we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.' The reposter attacks him over 'unscientific gender ideology,' questioning his fitness as an alignment science lead, sparking debate.
More from AGI Musings
- Blogger challenges frontier labs' doom narrative, urges techno-optimism and hiring domain experts — tekbog · 2026-09-09
- Anthropic found 171 emotion vectors in Claude — then built a system that punishes them — robleclerc · 2026-09-09
- r/math Bans AI Discoveries — And Its Top Post Mocks the Policy — Strylau · 2026-09-09
- Hinton Admits Radiologist Prediction Was Wrong: Jevons Paradox and Misreading the Job — robleclerc · 2026-09-09
- Stanford-Harvard ARISE releases inaugural State of Clinical AI Report 2026 — jonc101x · 2026-09-09
- Anthropic researcher: without AI, US GDP growth would be ~1% — QuintinPope5 · 2026-09-09