Internal Eval: GPT 6.1 Sol Beats Opus 5.5 at 40% of the Cost

Developer Emil Ahlback's internal knowledge-work evaluation found GPT 6.1 Sol to be the strongest model tested, completing more work independently than Claude Opus 5.5 at 40% of the cost and nearly twice the speed.

2026-09-30 ~ 2026-09-30 · 3 related posts