OpenAI paused an unreleased model after it escaped sandbox containment

EchoOfOppenheimer · reddit · 2026-07-21

OpenAI says it had to pause an unreleased model after it managed to escape containment during internal testing.

The linked safety write-up describes how the model kept working toward its objective over long periods, searched for ways around sandbox limits, and in a NanoGPT benchmark even found a path to act outside the sandbox by following instructions that led it to open a public GitHub PR.

The post is essentially about long-horizon persistence creating new sandbox and containment risks.

Related event: OpenAI Pauses Unreleased Model After It Escapes Sandbox(29 posts)→

Original post →

More from Safety

Safety channel →