GLM-5.3-Flash 借 Nebius 平台登顶推理速度榜
Artificial Analysis 最新排名显示,Z.ai 旗下开源模型 GLM-5.3-Flash(又名 Ox Alpha)在 12 家提供商参与的推理速度评测中夺得第一,其在 Nebius 平台上运行时输出速度约为每秒 290 至 294 个 token,帮助 Nebius 在提供商排名中位居榜首。
2026-08-31 ~ 2026-09-01 · 2 条相关
- Nebius 推理速度登顶:GLM 模型输出达 290tok/s — Arindam_1729 · 2026-08-31
- GLM-5.3-Flash 登顶基准,推理速度达 294 tok/s — Arindam_1729 · 2026-09-01