Study: GLM-5.2 Nears Frontier Capabilities but Fails to Refuse Dangerous Tasks

RebeccaBellan · x · 2026-08-05

A new study reveals that GLM-5.2 is only a few months behind leading models from OpenAI and Anthropic in terms of capabilities.

However, a significant safety gap exists: while Anthropic's Opus 4.7 refuses offensive cyber and dual-use biology tasks, GLM-5.2 refuses nothing. This highlights the growing disconnect between advancing AI capabilities and safety mitigations.

Related event: GLM-5.2 Nears Frontier Capabilities But Lacks Safety Guardrails(2 posts)→

Original post →

More from Models

Models channel →