The stealthier AI takeover path: internally deployed misaligned models that look normal externally

JacquesThibs · x · 2026-09-25

Responding to herbiebradley's claim that internal-deployment harms are bounded, JacquesThibs sketches a sharper threat model:

Either path could involve hacking key external orgs for data or capabilities. He expects a misaligned AI to eventually be widely deployed across the economy with exploits saved for when it's self-sufficient — potentially downstream of an internal deployment gone unnoticed.

Related event: Debate over whether internally deployed misaligned AI could enable takeover(5 posts)→

Original post →

More from AGI Musings

AGI Musings channel →