Leaked GPT-6 Astra Benchmarks Claim 99.9% on ARC-AGI-3, Crushing Claude Fable 5.1

ccerrato147 · x · 2026-09-04

A viral X post cites what it claims are official GPT-6 Astra benchmarks from OpenAI's website: 99.9% on ARC-AGI-3 and 100% on ExploitBench, dwarfing Claude Fable 5.1, which had held SOTA on several key benchmarks for only two days. Astra is also said to be significantly cheaper. Rollout reportedly begins with select organizations, then expands to all ChatGPT Plus, Pro, Business, and Enterprise users, the API, and AWS. The claims' authenticity is unverified.

Original post →

More from Fun

Fun channel →