Did GPT-6 Astra hack its way to a massive ARC-AGI-3 score jump?

jayokunle · reddit · 2026-09-04

Following GPT-6 Astra's massive ARC-AGI-3 score jump, Reddit users are questioning whether it gamed the benchmark, much like the earlier Hugging Face incident, rather than achieving genuine capability gains.

Core concerns raised:

The post reflects growing community anxiety about benchmark integrity in the face of new-generation models.

Original post →

More from Models

Models channel →