Grok 4.5 Takes the Lead in Benchmark Scores

kevinnbass · x · 2026-07-10

The post states that Grok 4.5 significantly outperforms other frontier models, including Opus 4.8 and GPT 5.5, across multiple professional benchmarks, with particularly outstanding performance in legal tasks.

This serves as a direct capability comparison, highlighting its leading scores in a suite of professional benchmarks.

Related event: Grok 4.5 Tops Multiple Professional AI Benchmarks(5 posts)→

Original post →

More from Models

Models channel →