Zhipu launches GLM-5.3-FlashX at up to 200 tokens/s, 2.5x Flash pricing
Zai_org · x · 2026-09-22
- Zhipu officially launched GLM-5.3-FlashX (model code: glm-5.3-flashx), with speeds up to 200 tokens/s.
- Priced at 2.5x GLM-5.3-Flash on both the Coding Plan and API.
- Available to all API users; Coding Plan subscribers can apply for trial access via a form.
Related event: Zhipu Launches GLM-5.3-FlashX at Up to 200 Tokens/s(2 posts)→
More from Models
- Audio8 ASR Infinite open-sourced: ultra-low latency 24/7 streaming ASR with no drift — TheMoonMidas · 2026-09-23
- 'Intelligence will get cheaper'? Devs mock rising AI subscription bills — zeeg · 2026-09-23
- Rumor: Tencent Hunyuan V4.1 may use HySparse2 architecture amid Chinese backbone wave — teortaxesTex · 2026-09-23
- Xiaomi's MiMo 2.6 Pro formalizes Li-Yorke chaos theorem in Lean 4 with 6,000+ verified lines — Dr_Singularity · 2026-09-23
- Testing whether Claude models notice when you swap models mid-conversation — Fable didn't refuse — davidmanheim · 2026-09-23
- Open benchmarks are broken: models regex out offloaded answers, researcher argues — a1zhang · 2026-09-23