Anthropic Alignment Lead Puts AI Extinction Risk Above 10%, Researcher Resigns Same Day

On September 9, two pieces of news about AI existential risk sparked intense discussion within and around Anthropic: Evan Hubinger, head of the alignment science team, publicly quantified his own estimate of extinction risk, while a pretraining researcher announced his resignation and slammed the frontier labs' race pace. Together, they once again exposed the deep tension inside frontier AI labs of "building while warning."

Confirmed

Unconfirmed

Why it matters

This is a rare quantified statement about extinction-level risk from a core figure at a frontier lab, and it dovetails with the same-day public departure of a safety-minded researcher—showing genuine internal disagreement at Anthropic over the pace of the AGI race and safety commitments, and reigniting outside debate over whether labs "say one thing and do another."

2026-09-09 ~ 2026-09-09 · 70 related posts

Primary sources

37 near-duplicate retellings: Miles_Brundage · peterwildeford · AGI Hunt · AICopyLab · Miles_Brundage · EvanHub · ramagetime · ccerrato147 · JosephJacks_ · AccBalanced · Polymarket · kevinnbass · AndyMasley · EvanHub · JeffLadish · sjgadler · JeffLadish · sahilypatel · sjgadler · SydSteyerhart · ctjlewis · birchlse · sjgadler · iamfakhrealam · builderjaydub · SIGKITTEN · builderjaydub · Delahuntagram · austinc3301 · austinc3301 · AGI Hunt · Tough_Control2052 · tomchapin · WonderFactory · rkulidzan · 新智元 · JFPuget