GPT-6 Sol Scores 1483 Elo on Knowledge Work, Far Behind Sonnet 5.5's 1811
haider1 · x · 2026-09-29
Blogger haider1 compiled long-horizon knowledge work results that look brutal for GPT-6 Sol:
- GPT-6 Sol: 1483 Elo, far behind the Claude lineup
- Claude Sonnet 5.5: 1811 Elo, effectively playing in Opus tier
- Claude Opus 5.5: 1822 Elo, still on top
Combined with his earlier Terminal-Bench findings — Sonnet 5.5 beats Opus 5.5 there and roughly ties it on knowledge work and computer use at half the price — Anthropic appears to have compressed most of the Opus experience into Sonnet this generation, while GPT-6 Sol lags badly on knowledge work.
More from Models
- Founder running 20+ startups says Opus 5.5 is AGI by his personal benchmarks — jonathan_wilke · 2026-09-29
- Opus 5.5 dramatically cuts em-dash usage, blurring AI writing detection — jonathan_wilke · 2026-09-29
- Sonnet 5.5 Sets Arena Record for Most Output Tokens, Sparking Pricing Debate — Gohab2001 · 2026-09-29
- PrivacyBench v2 launches: micro1's flow-transform 1.0 leads at 95.84%, 9.64 points ahead — Exp_Mark · 2026-09-29
- "Astra was incredible yesterday, terrible today": user reports overnight quality drop — HairyHobNob · 2026-09-29
- Grok 4.7 xHigh Tops Artificial Analysis Cyber Index for Enterprise Cyber Defense — XFreeze · 2026-09-29