Zhipu's GLM-5.3-Flash Lands on CoreWeave Serverless Inference

Zhipu's GLM-5.3-Flash is now available on CoreWeave's serverless inference, also accessible via Weights & Biases, offering 1M token context and vision support at $0.5 per million output tokens. It ranks among the top five open-weight models on Artificial Analysis's index with 18B active parameters.

2026-09-09 ~ 2026-09-10 · 2 related posts