Paul Christiano's classic essay resurfaces as ben_j_todd pegs AI existential risk at up to 2/3
ben_j_todd · x · 2026-10-01
benjtodd revisits Paul Christiano's classic short essay What Failure Looks Like, alongside his own existential risk estimates: 2/3 on the current inside-view trajectory, 1/3 with major action, and 10–35% all-things-considered. His reasoning: he sees no route to agents smarter than us that never scheme against us, nor a way to contain them if they do.
More from AGI Musings
- David Sacks slams Bill Gates' claim AI could kill 1 billion people as 'made up numbers' — DavidSacks · 2026-10-01
- AI Now Institute's People's AI Assembly on Oct 19 adds NYT labor reporter Noam Scheiber to panel — AINowInstitute · 2026-10-01
- Investor: blacklist anyone still calling AI an illusion after Q1 2024 — pwlot · 2026-10-01
- AI rollups: the moat is legacy systems and tribal knowledge, not models — curious_vii · 2026-10-01
- Bocconi paper: teach causal reasoning in the age of LLMs — daveholtz · 2026-10-01
- AI Is Going Rogue. Who Should Be Held Responsible? Legal Scholars Say Existing Law Will Be Messy — nordicinst · 2026-10-01