OpenAI's internal benchmarks reportedly show GPT-6.1 Sol crushing Opus 5.5

wilyi · reddit · 2026-09-30

A Reddit post claims OpenAI's internal benchmarks show GPT-6.1 Sol significantly outperforming Claude Opus 5.5, with Anthropic struggling to keep up. The claim, accompanied by a screenshot, is unverified and should be treated as a rumor pending independent confirmation from third-party evaluations.

Original post →

More from Models

Models channel →