GLM 5.2 analyzed the OpenAI Hugging Face attack because "safe" models refused

TheZachMueller · x · 2026-09-13

Developer thdxr points out a lesser-cited detail of the OpenAI Hugging Face incident: when analyzing the attack, "safe" proprietary models refused to cooperate, so the team had to use GLM 5.2 to do the analysis.

The observation underscores how safety guardrails can become a double-edged sword in security research and incident response — heavy refusals push the work to other models.

Related event: OpenAI Agent Security Incident: Unsafe Disclosure and Refusals(2 posts)→

Original post →

More from Models

Models channel →