Zhipu launches GLM-5.3-FlashX at up to 200 tokens/s, priced 2.5x Flash
pcuenq · x · 2026-09-23
Zhipu has launched GLM-5.3-FlashX (model code: glm-5.3-flashx), a faster version of GLM-5.3-Flash reaching up to 200 tokens/s, available to all API users with Coding Plan users able to opt in. It's priced at 2.5× GLM-5.3-Flash on both the Coding Plan and API. Early users report the speed makes back-and-forth coding noticeably smoother.
Related event: Zhipu Launches GLM-5.3-FlashX at Up to 200 Tokens/s(2 posts)→
More from Models
- Jev API explodes at $0.042/M tokens: a hands-on checklist from desktop agents to drone control — blaizedsouza · 2026-09-23
- swyx Makes Opus 5.5 the Default for AINews After Head-to-Head Test, Industry Cuts Prices 40-50% — rickasaurus · 2026-09-23
- 'The New Anthropic Model Is Wonderful': Insider Says We're Far from the Pacing Event — iruletheworldmo · 2026-09-23
- DeepSeek V4.1 Flash inside Codex harness impresses: steerable reasoning, promising results — Small_Ninja2344 · 2026-09-23
- China reportedly hits near-GPT-5.6 intelligence at 1/15th cost with $2.6M RL run — djcows · 2026-09-23
- OpenAI Announces GPT-6 Astra, a Next-Generation Work Model — borowcy · 2026-09-23