Opus 5 posts 30.2% on ARC-AGI-3 and leads several agentic benchmarks

socoolandawesome · reddit · 2026-07-25

Opus 5 posts a strong benchmark run

The image attached to the post shows Opus 5 ahead on several benchmarks, including:

The chart compares Opus 5 against Fable 5, Opus 4.8, and GPT-5.6 Sol, presenting Opus 5 as stronger on most listed tasks, especially agentic coding, search, and novel problem solving.

Related event: Anthropic Releases Claude Opus 5(40 posts)→

Original post →

More from Models

Models channel →