Opus 5 appears to lead a broad benchmark sweep across coding, search, and reasoning

signulll · x · 2026-07-25

Anthropic's Opus 5 is being compared against Fable 5, Opus 4.8, and GPT-5.6 Sol across a wide benchmark mix:

The image also shows Opus 5 ahead or near the top on several business and legal tasks, reinforcing the claim that it is a broad capability jump rather than a narrow coding win.

Related event: Anthropic Releases Claude Opus 5(40 posts)→

Original post →

More from Models

Models channel →