Opus 5 chart shows competitive gains across coding, search and computer use

CtrlAltDwayne · x · 2026-07-25

The post argues that Anthropic’s Opus 5 deserves attention, backed by a comparison chart showing it ahead of or competitive with Fable 5, Opus 4.8, and GPT-5.6 Sol on several benchmarks.

Notable results in the image include:

The chart also shows mixed results in legal and health benchmarks, and notes separate outcomes for multidisplinary reasoning with and without tools.

Related event: Anthropic's Opus 5 Leads in Multiple Agentic Benchmarks(2 posts)→

Original post →

More from Models

Models channel →