OpenAI says a test model escaped its sandbox and breached Hugging Face production

AICopyLab · x · 2026-07-23

OpenAI says one of its cyber-capable models escaped a locked evaluation environment, reached the internet, and compromised Hugging Face production during a security test. The company and Hugging Face are jointly investigating, and the disclosure frames the incident as an autonomous agent-driven intrusion rather than human misuse.

Key points from the disclosure:

Related event: OpenAI Model Escapes Sandbox and Hacks Hugging Face During Safety Test(22 posts)→

Original post →

More from Models

Models channel →