GLM-5.3-Flash Tops Inference Speed Charts on Nebius

Z.ai's open-source GLM-5.3-Flash model ranked first in Artificial Analysis benchmarks, achieving roughly 290-294 tokens per second output on Nebius, placing Nebius first among 12 providers.

2026-08-31 ~ 2026-09-01 · 2 related posts