GLM-5.3-Flash Released: 320B Total Params, Cost-Efficiency on Pareto Frontier
ArtificialAnlys · x · 2026-08-27
Zai has released GLM-5.3-Flash, a smaller and more cost-efficient sibling to GLM-5.3. It features 320B total parameters with 18B active parameters and supports low, high, and max reasoning efforts.
In Artificial Analysis benchmarks, GLM-5.3-Flash scores 57 on the Intelligence Index and sits comfortably on the Intelligence vs. Cost per Task Pareto frontier at $0.09 per task. It achieves a GDPval-AA v2 score of 1770, matching the GLM-5.3 Max. The evaluation used 149M output tokens, approximately 11% fewer than the GLM-5.3 model.
More from Models
- Rumor: Anthropic to launch Fable 5.1 before Astra is ready — haider1 · 2026-08-27
- Deepseek V4 Flash hits 420 tok/s in new community benchmark — HankYeomans · 2026-08-27
- Grok Bot gets more efficient with higher rate limits, users praise rapid improvement — XFreeze · 2026-08-27
- Users discuss stricter censorship in recent model updates — Connect-Cost-5504 · 2026-08-27
- Pokee-Isaac 28B Builds Playable Game in 5 Minutes with 10M Context — Kyrannio · 2026-08-27
- Goodfire AI Research: Efficiently Locating 'Forking Tokens' in LLMs — VoidAsuka · 2026-08-27