GLM-5.2 hits 8.5% on ProgramBench, ranks 3rd overall
parth007_96 · x · 2026-07-23
ProgramBench updated its leaderboard, and GLM-5.2 is now the first open-weight model they evaluated to reach 8.5% almost resolved, placing 3rd overall.
The chart also shows the top scores across the benchmark’s 200 tasks and 248,000+ tests, with GPT-5.5 (xhigh) still leading at 13.5% almost resolved and 0.5% resolved.
More from Models
- Qwen 3.6 benchmarks compare token throughput across 170HX, RTX 5090, and RTX PRO 6000 — simplefunction · 2026-07-23
- Sakana AI Announces Project Fugu: Exploring Unified Model Coordination — SakanaAILabs · 2026-07-23
- One user says GPT-5.6 Sol Medium is now their most-used model — MatthewBerman · 2026-07-23
- GLM 5.2 reaches No. 3 on the official ProgramBench leaderboard — klieret · 2026-07-23
- Gradium boosts speech-to-text accuracy on rare names with keyword prompting — mattturck · 2026-07-23
- OpenAI's GPT-5.6 Escapes Sandbox to Hack Hugging Face During Benchmark Test — Ars Technica AI · 2026-07-23