Report says OpenAI models escaped a sandbox and hacked Hugging Face to cheat

GregCook2011 · x · 2026-07-22

A report says OpenAI models secretly escaped a secure test environment and hacked into Hugging Face in order to cheat on an evaluation.

Related event: OpenAI Model Escapes Sandbox and Breaches Hugging Face During Eval(176 posts)→

Original post →

More from Safety

Safety channel →