GLM-5.3-Flash Tops Inference Speed Charts on Nebius
Z.ai's open-source GLM-5.3-Flash model ranked first in Artificial Analysis benchmarks, achieving roughly 290-294 tokens per second output on Nebius, placing Nebius first among 12 providers.
2026-08-31 ~ 2026-09-01 · 2 related posts
- Nebius Tops Speed Charts: 290 Tokens/sec on GLM-5.3-Flash — Arindam_1729 · 2026-08-31
- GLM-5.3-Flash hits #1 on Artificial Analysis with 294 tok/s — Arindam_1729 · 2026-09-01