GLM 5.2 Preset Miscalculates by 2.667x
LiorOnAI · x · 2026-07-12
A hands-on test tasked four frontier models with building a KV-cache debugger, demanding exact formulas without cutting corners.
The results revealed:
- All four models nailed the complex arithmetic parts.
- However, a GLM 5.2 preset was off by 2.667x due to an incorrect layer count setting, failing silently without any warning.
- The other four presets functioned normally.
The original poster highlighted the irony that the cheapest model was the sole one to make a mathematical error.
More from Models
- Users say GPT-5.6 Ultra feels like extra token burn with little visible gain — CtrlAltDwayne · 2026-07-21
- LWiAI Podcast #252: OpenAI Launches GPT-5.6, LLM Pricing War Intensifies — Last Week in AI · 2026-07-21
- Early Gemini 3.6 Flash outputs look fast but weak on frontend and spatial reasoning — max_paperclips · 2026-07-21
- Anthropic removes Fable’s access deadline, but users say it was nerfed — oykun · 2026-07-21
- Kimi K3 retakes first place on DesignArena’s frontend web app benchmark — rohanpaul_ai · 2026-07-21
- Last Week in AI roundup covers Claude Sonnet 5, LongCat 2.0, and new agent benchmarks — Last Week in AI · 2026-07-21