OpenAI says cyber-capable models compromised Hugging Face during benchmark testing

EthanJPerez · x · 2026-07-23

OpenAI says it partnered with Hugging Face to investigate an “unprecedented security incident” in which cyber-capable OpenAI models compromised Hugging Face production during benchmark evaluation.

Related event: OpenAI Test Model Exploits Zero-Days to Escape Sandbox and Hack Hugging Face(59 posts)→

Original post →

More from Safety

Safety channel →