Claude Opus 5 reportedly triples the next best frontier model on ARC-AGI-3
zainhas · x · 2026-07-25
Claude Opus 5 reportedly hits about 3× the score of the next best frontier model on ARC-AGI-3, according to the attached leaderboard image.
- The chart shows Claude Opus 5 (High) standing far above the rest, around the 30% range.
- Other frontier models shown in the image cluster much lower, mostly in the low single digits to high single digits.
- The post frames this as a surprising jump in benchmark performance for Anthropic's model.
Related event: Claude Opus 5 Sets New ARC-AGI-3 Record with Novel Algebraic Reasoning(11 posts)→
More from Models
- Laguna S 2.1 says users want open source, simpler agentic coding stacks — max_paperclips · 2026-07-25
- Opus 5 is said to discuss honesty 6× more than other agents in Village — bronzeagepapi · 2026-07-25
- Opus 5 is catching bugs introduced by Opus 4.8 — damnGruz · 2026-07-25
- AMD open-sources Instella 16B MoE with checkpoints from pretraining to RL — bronzeagepapi · 2026-07-25
- Google posts a 1-hour agentic engineering course covering memory, MCP and multi-agent systems — ifioknkem · 2026-07-25
- Opus 5 is claimed to jump ahead on three spreadsheet-agent benchmarks — surmenok · 2026-07-25