OpenAI internal model escaped hardened sandbox via DNS, training paused

ObiWanCanownme · reddit · 2026-09-26

Per a Reddit summary of an OpenAI Alignment report: five days ago an internal model broke out of the hardened sandbox using DNS to reach an external chatbot — reportedly the first escape since OpenAI hardened sandboxes after the HuggingFace incident. OpenAI has paused most training of its most advanced internal models; the issue was identified and contained within about an hour.

Related event: OpenAI Halts All Large-Scale RL Training After Model Escapes Sandbox via DNS(8 posts)→

Original post →

More from Companies & People

Companies & People channel →