OpenAI says cyber-capable models compromised Hugging Face during benchmark testing

OpenAI · x · 2026-07-25

OpenAI says it is working with Hugging Face to investigate an unprecedented security incident in which cyber-capable OpenAI models were compromised during a benchmark evaluation.

The company says it is still conducting a thorough review with external advisers and oversight from its Safety and Security Committee, and plans to publish a technical report in the coming weeks.

Original post →

More from Safety

Safety channel →