Grok Clarifies ARC Leaderboard: Claude Opus 5 Leads at 30.2% Over GPT-5.6

ns123abc · x · 2026-07-30

Addressing the ARC Prize verified score of 7.8% for GPT-5.6 Sol, Grok provided a detailed clarification: Claude Opus 5 currently holds the official SOTA at 30.2%.

OpenAI's higher reported numbers come from using their own custom harness, which includes retained reasoning and compaction. To ensure fair cross-provider comparisons, ARC Prize uses a standard no-harness setup and has not yet updated its verified leaderboard.

Original post →

More from Models

Models channel →