GLM-5.3-Flash hits CoreWeave: top-5 open model with just 18B active params
wandb · x · 2026-09-10
CoreWeave has deployed Z.ai's GLM-5.3-Flash on its Serverless Inference platform. The model ranks in the top 5 of 112 large open-weight models on the Artificial Analysis Intelligence Index with only 18B active parameters.
It supports 1M context and vision, priced at $0.15/M input and $0.50/M output tokens.
Related event: Zhipu's GLM-5.3-Flash Lands on CoreWeave Serverless Inference(2 posts)→
More from Models
- Codex CLI 0.154.0 ships GPT-6-Astra, experimental worktree support — github-actions[bot] · 2026-09-10
- OpenAI's new model, in training since Aug 28, reportedly beat GPT-6-Astra in just one week — alexcovo_eth · 2026-09-10
- Qwen3.8-2.4T-A95B open weights land on AWS: single 8×B300 node with vLLM — AWS ML Blog · 2026-09-10
- Users say Astra's $200 sub is no longer enough: multi-project work burns through quota in days — CtrlAltDwayne · 2026-09-10
- OpenAI launches GPT-6 Astra to power ChatGPT Work with desktop app control — OpenAI · 2026-09-10
- Intelligence Index v4.3: Claude Fable 5.1, Muse Spark 1.3 and GPT-6 Astra Reset the Cost-Efficiency Frontier — ArtificialAnlys · 2026-09-10