Anthropic eval: open GLM-5.3 nears Claude in exploit capability with weak safeguards
kimmonismus · x · 2026-09-30
- Anthropic published an evaluation finding that Zai's openly downloadable GLM-5.3 approaches Claude Mythos Preview in exploit capabilities, with safeguards that are easy to bypass.
- On ExploitBench, GLM-5.3 built working browser exploits in 50 of 410 attempts vs 56 for Mythos Preview.
- In a separate controlled experiment, researchers used GLM-5.3 to discover previously unknown vulnerabilities, raising questions about open-weights frontier model safety.
More from Models
- OpenAI's GPT-6.1 Sol system card rates the model Critical in cybersecurity — maksym_andr · 2026-09-30
- 5 months after Mythos Preview panic, an open model already matches it — mariofilhoml · 2026-09-30
- GPT 6.1 Sol Only +3 on BridgeBench, 82 Points Behind Astra: Benchmaxing Suspected — RexDouglass · 2026-09-30
- Anthropic's Latest Blog Post Mentions Zhipu's GLM 5.3 — gnukeith · 2026-09-30
- GPT-6.1 Sol beats 2x-cost models on ClickUp's knowledge-work benchmark — mathemagic1an · 2026-09-30
- User: Upgraded to OpenAI's $500 plan, got silently downgraded to cheaper models — kieranklaassen · 2026-09-30