Anthropic Alignment Lead Backs 10% Extinction Risk Claim
Former Anthropic employee Jacob Coxon claimed AI has a greater than 10% chance of killing all humans within a decade, a view Anthropic alignment lead Evan Hubinger endorsed, drawing widespread criticism. Andy Hall pushed back in a Substack essay rejecting near-term extinction narratives.
2026-09-13 ~ 2026-09-14 · 2 related posts
- Episode 1: Ex-Anthropic/OpenAI researcher warns AI may escape human control(2026-09-12, 2 posts)
- Episode 2: Ex-OpenAI/Anthropic researcher Jacob Coxon resigns publicly, warning of superintelligence gamble(2026-09-13, 8 posts)
- Episode 3: Anthropic Alignment Lead Backs 10% Extinction Risk Claim(2026-09-13, 2 posts)
- Andy Hall: Why He Doesn't Buy Imminent Extinction, as Anthropic Exits and 10%+ Warnings Mount — soumitrashukla9 · 2026-09-13
- Anthropic alignment lead agrees AI "kill all humans" odds exceed 10% in next decade — firstadopter · 2026-09-14