GLM-5.3 Cyber Scores Surge: Same Base Model, Targeted Post-Training

ChrisGPT · x · 2026-08-14

The author points out that GLM-5.3 uses the exact same base model as GLM-5.2, but targeted cyber post-training led to massive leaps in its cybersecurity benchmark scores.

Specifically, CyberGym scores jumped from 77.2 to 84.5, ExploitBench from 24.4 to 54.4, and completed ExploitGym tasks surged from 29/39 to 105/130 at the 2-hour and 6-hour marks. The author argues that training for a specific domain isn't cheating, unlike training on exact test cases. Additionally, web access was blocked and Git metadata removed during evaluation, though no contamination audit was disclosed.

Related event: Zhipu AI Releases GLM-5.3 with Boosted Coding and Cybersecurity, Open Weights in Two Weeks(16 posts)→

Original post →

More from Models

Models channel →