GLM-5.3-Flash Review: 10% Cost, Pareto Frontier Performance
ArtificialAnlys · x · 2026-08-27
Zai released GLM-5.3-Flash, a 320B total / 18B active MoE model under the MIT license. Benchmarks: Scores 57 on the Artificial Analysis Intelligence Index, only 3 points behind GLM-5.3 and tying GPT-5.6 Terra; matches GLM-5.3 on real-world agentic tasks like GDPval-AA v2 and Terminal-Bench v2.1. Cost: Cost per Task is $0.09 (1/7.5th of GLM-5.3's $0.68). API pricing is $0.15/1M input and $0.50/1M output tokens, with 80% discount on cached inputs. Features: Supports low/high/max reasoning efforts, 400k context window, and improved hallucination rate (28% vs 30%).
Related event: GLM-5.3-Flash Released Open-Source: 10% Cost, Top Value Frontier Model(3 posts)→
More from Models
- Grok Bot gets more efficient with higher rate limits, users praise rapid improvement — XFreeze · 2026-08-27
- Pokee-Isaac 28B Builds Playable Game in 5 Minutes with 10M Context — Kyrannio · 2026-08-27
- Goodfire AI Research: Efficiently Locating 'Forking Tokens' in LLMs — VoidAsuka · 2026-08-27
- Expert: Inkling's Benchmarks Didn't Translate to Good UX — ethayarajh · 2026-08-27
- India's Sarvam: 105B-parameter homegrown LLM shines in Hindi testing — abhish18 · 2026-08-27
- Sarvam AI's speech-to-text and foundation models show maturity — abhish18 · 2026-08-27