Anthropic report flags GLM-5.3 as strongest open model for cyber offense, drawing self-serving criticism
oran_ge · x · 2026-09-30
After OpenAI added GLM 5.3 and other open models to its enterprise tier, Anthropic published a report criticizing open-model safety, drawing widespread skepticism.
Report highlights:
- Calls GLM-5.3 "the strongest open model for cybersecurity capabilities to date," roughly 4 months behind US frontier models.
- Among models tested (Opus 4.6, Kimi K3, DeepSeek Flash), GLM 5.3 was the only one with exploit capabilities.
- Argues open weights are inherently unsafe: abliteration can strip refusal behavior for $4,400, and jailbreak-style prompts or prefill thinking can bypass guardrails; Claude is "safe" because it's closed and offers no prefill.
- Concludes enterprises should urgently adopt safe closed models.
Author's critique: the report is self-serving security theater — removing a model's guardrails to prove it's unsafe is circular; if Anthropic cares about open-model safety it could contribute to it, rather than moralizing while banning accounts.
More from Models
- ChatGPT Pro's $200 plan reportedly includes 62,500 Codex credits expiring Dec 31 — chaumian · 2026-09-30
- User calculates 62,500 credits ≈ $2,500 of GPT-6.1 API usage, calling the new plan a big cut — chaumian · 2026-09-30
- Opus 5.5 keeps flagging mundane Claude Code sessions as [cyber], downgrading users to 4.8 — jonntanny · 2026-09-30
- Reddit user reports surprise 62,500 credit grant, about 15x their monthly plan allowance — IronDarbe · 2026-09-30
- Users question why ChatGPT chat mode still misses the newest GPT models — koltregaskes · 2026-09-30
- User receives 62,500 credits worth ~$2,500, equal to 12.5 months of the $200 Pro plan — kimmonismus · 2026-09-30