Opus 5.5 Sweeps GPT-6 Sol on Benchmarks, but Sol Runs Tasks at Half the Cost
Vast-Grapefruits · reddit · 2026-09-23
With Anthropic and OpenAI shipping Opus 5.5 and GPT-6 Sol the same day, one Redditor compiled a full head-to-head: Opus wins every benchmark, and Sol actually scores 100 Elo lower on GDPval than the GPT-5.6 Sol it replaces — mostly from weaker deliverables, not reasoning. Sol's edge is cost: half the per-token price, 31K output tokens per task vs Opus's 119K at max effort, running the whole Artificial Analysis index for $1.06 per task.
Related event: Claude Opus 5.5 Launches, Tops Intelligence Index While Cutting Prices(50 posts)→
More from Models
- Third-party test: Claude Opus 5.5 renders finer 3D scenes but costs 13x more than GPT-6 Sol — testingcatalog · 2026-09-23
- Tester claims Claude Opus 5.5 has the best visual design output of any model tested — burny_tech · 2026-09-23
- GPT-6 Sol Codex system prompt leaked: over 294,000 characters dumped on GitHub — gaganghotra_ · 2026-09-23
- Claude 5.5 (live) keeps generating user turns, reports user — BlackHC · 2026-09-23
- Code benchmarks are mostly slop: dev calls for narrow evals per domain, not one score — almmaasoglu · 2026-09-23
- Grok 4.7 turns an envelope sketch into playable puzzle game Lightweave in minutes — luismbat · 2026-09-23