Opus 5 reaches 30.2% on ARC-AGI 3 as critics question the benchmark
ChrSzegedy · x · 2026-07-25
Opus 5 scores 30.2% on ARC-AGI 3, as ARC-AGI gets dismissed as an AGI proxy
Chris Szegedy says ARC-AGI has nothing to do with AGI, in response to a post claiming “AGI is near.”
- The thread cites Opus 5 scoring 30.2% on ARC-AGI 3.
- The argument is less about the score itself and more about whether ARC-style benchmarks meaningfully measure AGI.
- It is a concise data point plus a broader benchmark-validity critique.
Related event: Opus 5's High ARC-AGI-3 Score Sparks Cheating and Overfitting Controversy(11 posts)→
More from AGI Musings
- Mathematician Daniel Litt Launches Problem Repo to Track Human vs AI Progress: 15 Problems, 1 Solved — littmath · 2026-09-11
- Should AI models be taught morality? Breakout incidents expose missing ethical training — Pfungus_ · 2026-09-11
- SoftBank's Masayoshi Son predicts 100 trillion self-replicating AIs: "humans' era as top life form is ending" — Puzzleheaded-King584 · 2026-09-11
- We are witnessing the unreasonable effectiveness of inference-time scaling — sqcai · 2026-09-11
- The AlphaFold lesson: AI-solved math may mean fewer mathematicians needed — kiki-le-koala · 2026-09-11
- Accelerationist fires back at AI doomers: beliefs aren't arguments — Dan_Jeffries1 · 2026-09-11