GLM-5.3 shows strong cost-efficiency vs GPT 5.6 on TerminalBench-3.0
zainhas · x · 2026-08-25
GLM-5.3 demonstrates competitive performance on the TerminalBench-3.0 benchmark. It achieves a 32% success rate at $24 per task, compared to GPT 5.6 Sol's 35% at $54, while GPT 5.6 Luna trails significantly at 14% for $21.5.
More from Models
- Users complain GPT-5.6 Sol over-engineers tasks with verbose code — Lenox_Shawn · 2026-08-25
- Local LLM Test: Qwen2.5 72B Hits 100 tok/s on RTX 4090 — julianharris · 2026-08-25
- Grok Voice tops Speech-to-Speech Index, outperforming all GPT Realtime models — XFreeze · 2026-08-25
- Redditor finds MiniMax H3 performs far better with Mandarin prompts than English — apoke890 · 2026-08-25
- Suspected Gemini 3.8 Flash leak surfaces on Reddit — Last_Conclusion_8984 · 2026-08-25
- Train Only Projector to Add New Modalities Without Regressing LLM Capabilities — rohanpaul_ai · 2026-08-25