Terminal-Bench 4.0: GLM-5.3 Surpasses GPT-5.6 in New Ranking

eyishazyer · x · 2026-08-29

The new Terminal-Bench 4.0 leaderboard delivered surprising results, with scores dropping across the board due to increased difficulty, moving away from the clustering seen in previous versions.

This shift indicates that harder benchmarks provide better differentiation of actual model capabilities.

Original post →

More from Models

Models channel →