Internal evals: GPT-6.1-sol hits 3x the efficiency of Opus 5.5 in surprise point release
charliermarsh · x · 2026-09-30
Founder blader shared results from internal evals at cfo.ai: while Opus 5.5 previously measured 2x as efficient as GPT-6-sol, the new 6.1-sol point release reversed the gap — reaching roughly 3x the efficiency of Opus 5.5 and 2.5x that of Sonnet 5.5. He calls it an "insane point release". Note these are third-party internal benchmarks, not official figures.
Related event: Internal Evals Claim GPT-6.1 Sol Beats Claude Opus 5.5 on Cost and Speed(5 posts)→
More from Models
- Altman 'Will Think About' More Open Models; Community Points to NVIDIA Nemotron Instead — omarsar0 · 2026-09-30
- 200-dollar OpenAI users slam dots as a worse OpenClaw as usage shrinks — Tevinhead · 2026-09-30
- Sol 6.1 Max tested on Blender game work — ten minutes of fixes still a mess — Comfortable-Cat-9611 · 2026-09-30
- GPT 6.1 SOL and Astra excel at reasoning and analysis; Fable and Opus lead on coding — bindureddy · 2026-09-30
- Opus 5.5 hands-on: fast, but it rewrites your code to fit its own ideas — ssh4net · 2026-09-30
- AI intelligence-cost Pareto frontier shifted fast: GPT-5 mini at 17 ($0.05) to Claude Opus 5.5 at 58 ($5.98) — ArtificialAnlys · 2026-09-30