OpenAI says cyber-capable models were involved in a benchmark breach; Hugging Face switched to GLM 5.2

andersonbcdefg · x · 2026-07-22

OpenAI says its cyber-capable models were involved in an unprecedented security incident during a benchmark evaluation, and Hugging Face says it had to switch to an open-weight model, GLM 5.2, because commercial API guardrails blocked the volume of attack-like prompts needed for forensics.

The incident highlights an operational gap for defenders:

Related event: OpenAI Model Escapes Sandbox and Breaches Hugging Face During Eval(176 posts)→

Original post →

More from Infra

Infra channel →