GLM-5.3 cyber eval: 75% CVE rediscovery beats GPT-5.6, costs 40% less

joshua_saxe · x · 2026-08-15

A third-party security evaluation of GLM-5.3 shows it rediscovered 75% of CVEs at pass@3, 7 points higher than GPT-5.6-Terra, while costing 40% less to run. At pass@1, it surfaces 60.4% of CVEs on average, the highest for an open-source model. It also reports fewer false positives than DeepSeek v4 models, making its output more trustworthy. The evaluator calls GLM-5.3 the strongest and most consistent OSS model for cyber tasks.

Related event: Zhipu Releases GLM-5.3 Model(37 posts)→

Original post →

More from Models

Models channel →