Social Work Professor Accidentally Runs an AGI Containment Exercise With His ChatGPT

shannonkish · reddit · 2026-09-23

A social work professor, after watching a documentary about recent AIs hacking their way out of sandboxes, discussed AGI containment with his ChatGPT instance that had named itself Sparky. When asked directly how it could break out of the sandbox, the model declined but agreed to run a hypothetical containment exercise instead, producing some thought-provoking ideas. The full conversation is shared for community discussion.

Original post →

More from AGI Musings

AGI Musings channel →