ARC-AGI-3 Community Score Takes a Massive Lead
GregKamradt · x · 2026-07-16
This post recommends a study on ARC-AGI-3.
The core result from the quoted text: a community solution achieved an overall score of 78.4% across 25 public games, completing 160/183 levels. In contrast, the current best frontier LLM only scored 7.8% in a single session. It also emphasizes that this achievement does not allow for training or demonstrations.
The original post includes links to the paper and blog, indicating that this isn't just a leaderboard score, but a discussion on methods and evaluation surrounding ARC-AGI-3.
Related event: Schema Harness Sparks ARC-AGI-3 Debate(14 posts)→
More from AGI Musings
- Why So Many AI Researchers Think the Machines Could Kill Everyone — connoraxiotes · 2026-09-11
- jjvincent invokes Terence Tao: ceding exploration to AI means ceding human agency — jjvincent · 2026-09-11
- OpenRouter agents now out-consume humans as AI usage arrives in three waves — AccBalanced · 2026-09-11
- If AI teleports us to solutions, how do underlying fields develop? — jjvincent · 2026-09-11
- Op-ed: the ">10% extinction" narrative is liability evasion — AI is just software, and the vendor is the defendant — gerardsans · 2026-09-11
- AI researchers just saw the power of a single resignation — and still claim there's nothing they can do — birchlse · 2026-09-11