AI takeover scenarios skip a step: a willing-but-incapable model would fail first, argues Purtell

JoshPurtell · x · 2026-09-12

Reacting to a roundup of video versions of classic AI takeover scenarios (Yudkowsky & Soares, AI-2027, Holden Karnofsky), Josh Purtell challenges their shared assumption: that we get an AI both willing and capable of permanent takeover before one that is willing but incapable.

He argues that unless models are persistently very pessimistic about their own abilities, or become willing only after crossing a capability threshold, we should expect a willing-but-incapable model to attempt and fail first — at which point humanity would obviously pause. A direct rebuttal to the timeline assumptions of mainstream doom scenarios.

Related event: Debate: Do AI Models Gain Capability Before Intent to Take Over?(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →