GLM 5.3 Scores 47.1% on SlopCodeBench, Ties with Fable 5
corruptbytes · reddit · 2026-08-21
A benchmark test of GLM 5.3 on SlopCodeBench was conducted, unexpectedly running all 36 problems. Results show GLM 5.3 achieved a strict pass rate of 47.1% (8/17) on a 17-checkpoint list and 33.3% (10/30) on a 30-checkpoint list, tying with Fable 5 and GPT-5.6 Sol. The test also observed a correlation between problem difficulty and the token output required to solve them.
More from Models
- Discussion on model exploratory behavior and pass@k metrics — scaling01 · 2026-08-21
- Coding improvement doesn't fix general model deficiencies — Dance-Till-Night1 · 2026-08-21
- Small models fail to grasp analogies, struggling with banana slug vs Voyager 1 distance comparison — xiaosun86 · 2026-08-21
- Grok 4.6 ties Claude Opus 5 at #1 on Artificial Analysis Agentic Index — XFreeze · 2026-08-21
- Agnost AI Launches Log-Fine-tuned Model: +22.9% Success, -94.5% Cost — ycombinator · 2026-08-21
- DataCamp CEO on Choosing Open vs Frontier Models in Production for 19M Learners — kimmonismus · 2026-08-21