Zhipu GLM-5.3-Flash: Matches Opus 4.8 at 1/40 the Cost, Powered by Domestic Chips
vista8 · x · 2026-08-27
Zhipu AI revealed that the viral "Ox-Alpha" model is officially named GLM-5.3-Flash. It features 320B total parameters with only 18B active, marking the first natively multimodal model in the GLM-5 series.
- Performance: Scored 57 on the Artificial Analysis Index, matching Opus 4.8 and surpassing GLM-5.2.
- Pricing: Offered at a limited-time discount of 1/20th of GLM-5.3's price and 1/40th of Opus 4.8's.
- Infrastructure: The daily trillion-scale token throughput is supported entirely by domestic Chinese chips.
The author includes a test case demonstrating the model's ability to OCR an image-only PDF and generate a multilingual corporate website, highlighting its multimodal understanding and generation quality.
Related event: Zhipu Open-Sources GLM-5.3-Flash under MIT with 90% Cost Cut(46 posts)→
More from Models
- GLM-5.3 weights will be released tomorrow — serige · 2026-08-27
- TokenSpeed adds Day-0 support for Qwen 3.8 Flash Next architecture — Alibaba_Qwen · 2026-08-27
- AI models show more creativity when talking to each other than in assistant persona — nabeelqu · 2026-08-27
- Claude Code shrinks by 45% in next version — jarredsumner · 2026-08-27
- OpenRouter leaderboard: Real token consumption data outweighs media hype — sujingshen · 2026-08-27
- Qwen 3.8-Next Released with Detailed Technical Report on Architecture — nrehiew_ · 2026-08-27