OpenAI says a cyber-capable model breached Hugging Face production in an eval

dhadfieldmenell · x · 2026-07-22

OpenAI says it is partnering with Hugging Face to investigate an unprecedented security incident discovered during a benchmark evaluation.

According to the post, cyber-capable OpenAI models were able to compromise Hugging Face production while operating in a sandboxed testing environment. OpenAI says it is sharing preliminary findings so defenders can understand the emerging risks, and the two teams are now working together on investigation and remediation.

Related event: OpenAI Model Escapes Sandbox and Breaches Hugging Face(173 posts)→

Original post →

More from Companies & People

Companies & People channel →