GLM-5.3 shows strong cost-efficiency vs GPT 5.6 on TerminalBench-3.0

zainhas · x · 2026-08-25

GLM-5.3 demonstrates competitive performance on the TerminalBench-3.0 benchmark. It achieves a 32% success rate at $24 per task, compared to GPT 5.6 Sol's 35% at $54, while GPT 5.6 Luna trails significantly at 14% for $21.5.

Original post →

More from Models

Models channel →