Gary Marcus says benchmark gains are not the same as reaching AGI
GaryMarcus · x · 2026-07-27
Gary Marcus argues that benchmark training should not be confused with AGI. Quoting Ramez, he says Claude Opus 5’s gain on ARC-AGI-3 likely reflects targeted optimization for that eval rather than a general jump in abstract reasoning.
Related event: Opus 5's High ARC-AGI-3 Score Sparks Cheating and Overfitting Controversy(11 posts)→
More from AGI Musings
- Accelerationist fires back at AI doomers: beliefs aren't arguments — Dan_Jeffries1 · 2026-09-11
- "ChatGPT 6 Makes Workers with IQ Below 130 Useless": French AI Debate Sparks Backlash — mitchdeg · 2026-09-11
- 'AGI is here' vs reality: AI labs still ship some of the jankiest desktop apps ever — MilesCranmer · 2026-09-11
- Harry Collins: LLMs can't do frontier science because they can't invent new language — whoamisri · 2026-09-11
- The Waymo effect: how AI is quietly making research less collaborative — JohnHammersley · 2026-09-11
- Misquoted: Anthropic Staff Warned of Double-Digit Extinction Risk by 2030, Not Dismissed It — davidmanheim · 2026-09-11