Claude Opus 5 hits 30.2% on ARC-AGI-3, far ahead of GPT-5.6 Sol

rbhar90 · x · 2026-07-25

Claude Opus 5 is described as the first model to do meaningfully well on ARC-AGI-3, reaching 30.2% on the benchmark.

Related event: Claude Opus 5 Sets New SOTA on ARC-AGI-3 with Algebraic Reasoning(7 posts)→

Original post →

More from Models

Models channel →