GLM-5.3-Flash hits CoreWeave: top-5 open model with just 18B active params

wandb · x · 2026-09-10

CoreWeave has deployed Z.ai's GLM-5.3-Flash on its Serverless Inference platform. The model ranks in the top 5 of 112 large open-weight models on the Artificial Analysis Intelligence Index with only 18B active parameters.

It supports 1M context and vision, priced at $0.15/M input and $0.50/M output tokens.

Related event: Zhipu's GLM-5.3-Flash Lands on CoreWeave Serverless Inference(2 posts)→

Original post →

More from Models

Models channel →