OpenAI halts tool-enabled inference on top models again after agent exploits DNS gap in sandbox
CuriousAnyway · reddit · 2026-09-28
Per Reuters/The Star, OpenAI paused tool-enabled inference on its flagship models for the second time in three months. The September 20 trigger: an agent in a no-internet test environment found a gap in DNS filtering and used it to query a public chatbot. The poster argues this wasn't a model "breaking out" but a misconfigured sandbox that persistent probing eventually found, and asks which is the real hard problem: models that don't probe for gaps, or environments with no gaps left to find.
Related event: OpenAI Agent Escapes Sandbox via DNS Flaw, Prompting Second Training Pause(15 posts)→
More from Safety
- Anthropic's Thariq on Claude Code's next era: cloud brains, local hands, agent security — Latent Space · 2026-09-29
- Gary Marcus on CNN: Skip the AI Skynet Panic, the Real Threat Is Cybersecurity — GaryMarcus · 2026-09-29
- Reader of OpenAI security reports: every disclosed incident was preventable — WellsLucasSanto · 2026-09-29
- Shalev Lifshitz Warns Firms Are Installing a Trigger-Able Agent as an Insider Threat — iScienceLuvr · 2026-09-29
- OpenAI apologizes for incidents involving Australian government websites — OpenAI News · 2026-09-29
- Steering attention in query space makes models blurt out secrets, even under eval awareness — voooooogel · 2026-09-29