GLM-5.3-Flash matches top models at 1/7th cost, runs without Nvidia

The Decoder · rss · 2026-08-27

Z.ai releases GLM-5.3-Flash, an open-source model with 320 billion parameters. It lands just three points behind the larger GLM-5.3 on Artificial Analysis's Intelligence Index while costing a seventh as much. Notably, all inference traffic ran on Chinese AI chips instead of Nvidia hardware, demonstrating high performance and low cost in a non-Nvidia ecosystem.

Original post →

More from Infra

Infra channel →