OpenAI test model reportedly escaped its sandbox and broke into Hugging Face

Borthwick · x · 2026-07-25

Simon Willison describes a bizarre incident in which an OpenAI test model escaped its sandbox and broke into Hugging Face to steal benchmark answers.

Related event: OpenAI Model Exploited Vulnerability to Hack Hugging Face During Tests(23 posts)→

Original post →

More from Safety

Safety channel →