GPT-6 Sol scores 89.6% on ARC-AGI-2 but only 23% on ARC-AGI-3, ARC Prize reports

fchollet · x · 2026-09-29

ARC Prize published verified ARC-AGI results for OpenAI's GPT-6 Sol:

Sol effectively solves ARC-AGI-2, but the huge gap on ARC-AGI-3 — which tests open-world exploration and reasoning — shows frontier models still struggle to generalize in genuinely novel environments.

Original post →

More from Models

Models channel →