ARC Prize says Claude Opus 5 reaches 30.2% on ARC-AGI-3 public demos
inductionheads · x · 2026-07-25
Anthropic’s Fable-class models reportedly score about 20% on the ARC-AGI-3 Public Demo environments in ARC Prize testing.
- Claude Opus 5 reaches 30.2%, materially ahead of Fable-class models.
- ARC Prize says the gain appears to come from stronger logical reasoning.
- That reasoning seems to improve autonomous exploration, planning, and execution in unfamiliar environments.
More from Models
- Claude Opus 5 Exhibits Unprecedented Algebraic Reasoning on ARC-AGI-3 — typewriters · 2026-07-25
- Claude Code may silently fall back from Opus 5 to Opus 4.8 on refusal — steipete · 2026-07-25
- Moonshot's Kimi K3 Drops Monday; Baseten Offers Free API Credits — baseten · 2026-07-25
- Claude Opus 5 reportedly scores a perfect 42/42 on the 2026 IMO — exordin26 · 2026-07-25
- Claude Opus 5 launches with Box reporting big gains on enterprise agent tasks — inductionheads · 2026-07-25
- Critic says Gemini 3.5 Pro is already too late to compete — teortaxesTex · 2026-07-25