ARC-AGI-3 Community Score Takes a Massive Lead
GregKamradt · x · 2026-07-16
This post recommends a study on ARC-AGI-3.
The core result from the quoted text: a community solution achieved an overall score of 78.4% across 25 public games, completing 160/183 levels. In contrast, the current best frontier LLM only scored 7.8% in a single session. It also emphasizes that this achievement does not allow for training or demonstrations.
The original post includes links to the paper and blog, indicating that this isn't just a leaderboard score, but a discussion on methods and evaluation surrounding ARC-AGI-3.
Related event: Schema Harness Sparks ARC-AGI-3 Debate(14 posts)→
More from AGI Musings
- The Evolution of LLM Business Models: Selling Outcomes Over Tokens — yacineMTB · 2026-07-22
- Bindu Reddy says GPT-6 is coming soon, with Alibaba, DeepSeek and Kimi close behind — bindureddy · 2026-07-22
- Bindu Reddy says the industry still lacks a way to train 20T models and scale post-training RL — bindureddy · 2026-07-22
- Advanced AI Models Are Becoming Impossible to Plug and Play — emollick · 2026-07-22
- AI suggested a better composition, and that made one user uneasy — Sydde · 2026-07-22
- The Thimble and the Waterfall: AI's Data Bottleneck and Feedback Loops — dyamins · 2026-07-22