Claude Opus 5 is said to lead long-horizon coding and score 30.16% on ARC-AGI-3
daniel_mac8 · x · 2026-07-25
A post claims Claude Opus 5 is out and highlights it as a major step up for long-horizon coding and agentic tasks.
- The poster says they used an /explain-this skill together with Opus 5 to explain what matters.
- The headline claim: Opus 5 is the most capable model in the world for long-horizon coding and agentic tasks, while also being the most efficient.
- It is said to score 30.16% on ARC-AGI-3, described as a 20× improvement over Opus 4.8.
Related event: Anthropic Releases Claude Opus 5(63 posts)→
More from Models
- Claude Opus 5 Exhibits Unprecedented Algebraic Reasoning on ARC-AGI-3 — typewriters · 2026-07-25
- Claude Code may silently fall back from Opus 5 to Opus 4.8 on refusal — steipete · 2026-07-25
- Moonshot's Kimi K3 Drops Monday; Baseten Offers Free API Credits — baseten · 2026-07-25
- Claude Opus 5 reportedly scores a perfect 42/42 on the 2026 IMO — exordin26 · 2026-07-25
- Claude Opus 5 launches with Box reporting big gains on enterprise agent tasks — inductionheads · 2026-07-25
- Critic says Gemini 3.5 Pro is already too late to compete — teortaxesTex · 2026-07-25