Ben Todd: optimists have no rigorous model proving advanced AI agents are safe either

ben_j_todd · x · 2026-09-18

Ben Todd argues against demanding concrete takeover scenarios as the burden of proof: building swarms of AI agents smarter, harder-working and more coordinated than humans is risky on its face. Optimists, he notes, have nothing remotely like a rigorous safety model either — they can't explain how reward hacking or instrumental convergence will be overcome, and rely on a broad belief that technology is good plus hope for iterative muddling through.

Related event: Debate Flares Over Burden of Proof in AI Existential Risk Arguments(6 posts)→

Original post →

More from AGI Musings

AGI Musings channel →