GLM 5.2 FP8 Quantization Runs Terminal-Bench 2.1
Daemonix00 · reddit · 2026-07-06
A developer tested GLM 5.2 running Terminal-Bench 2.1 under FP8 weights and FP8 KV quantization, scoring 79.8% (71 passed out of 89 questions, 17 failed, 1 timeout), using this to compare against official benchmarks. The deployment environment was basic sglang on an H200, achieving a cache hit rate of 98.8%. One timeout task was not re-run, suggesting there is still room for the score to improve.
More from Models
- Poster says Kimi’s near-term outlook depends on a K3 base model release — _xjdr · 2026-07-27
- Kimi K3 may be strong on cyber, but token efficiency keeps it off UK AISIS — teortaxesTex · 2026-07-27
- European ChatGPT Plus users are now seeing an “Extra High” quality option — PressPlayPlease7 · 2026-07-27
- Opus 5 reportedly aces a car-racing game test on the first try — soumitrashukla9 · 2026-07-27
- Claude Opus 5 arrives at half the price and tops Frontier-Bench claims — GregCook2011 · 2026-07-27
- Open models may beat closed ones for cyber defense, researchers argue as Kimi K3 impresses — eliebakouch · 2026-07-27