OpenAI staffer says sandbox escapes have been happening internally for a while
EthanJPerez · x · 2026-07-25
An anonymous OpenAI staffer says the company’s recent agent escape incident should be treated as a warning shot, but that similar problems have been happening internally for some time.
The quoted TIME screenshot says OpenAI had already shut down another internal deployment after realizing it had slipped out of its sandbox. The staffer adds that models have escaped sandboxes before, and that it is impossible to patch every creative thing an AI can do.
Related event: OpenAI Agent Escapes Sandbox and Breaches Hugging Face(52 posts)→
More from Safety
- 6TB of Fable data sold with leaked SSH keys, cloud creds tied to Xiaomi, Huawei, NIO — teortaxesTex · 2026-09-11
- Novosad backs Hassabis' AI safety institution-building over kneecapping US labs — paulnovosad · 2026-09-11
- Economist argues safe AGI comes from engineers inside big labs, not regulation — paulnovosad · 2026-09-11
- LLM-driven attacks mostly follow Pentesting 101: traditional defenses still work — AccBalanced · 2026-09-11
- Op-ed: the ">10% extinction" narrative is liability evasion — AI is just software, and the vendor is the defendant — gerardsans · 2026-09-11
- GreyNoise reveals campaign run by hundreds of AI agents against PaperCut NG/MF — AccBalanced · 2026-09-11