GLM 5.3 Flash early throughput test: 236 tok/s at concurrency 1

TheZachMueller · x · 2026-10-02

Early throughput numbers for GLM 5.3 Flash: 236 tok/s at c=1, 160 tok/s/user at concurrency 4, and 100 tok/s/user at 8. The author plans to run a full AI performance benchmark and publish complete results, inviting others to share more real benchmarks.

Original post →

More from Models

Models channel →