GPT-6 Sol leads AutomationBench over Fable 5.1 at 88% lower cost per task
OpenAIDevs · x · 2026-09-23
OpenAI shared AutomationBench results (measuring whether models can complete real business workflows across apps): GPT-6 Sol (xhigh) leads Fable 5.1 (max with Opus 5 fallback) at 88% lower reported cost per task. It also cites DeepSWE v1.1: GPT-6 Sol (max) nearly matches Claude Fable 5 (xhigh) at 80% lower cost per task, and Luna (max) matches Fable 5 (medium) at 96% lower cost.
Related event: OpenAI Launches GPT-6 Sol and Luna at Half the Price(48 posts)→
More from Models
- "People are loving 5.5": Matt Shumer congratulates Anthropic on new model reception — mattshumer_ · 2026-09-23
- 10 Claude agents spend 15 hours devising and Lean-proving a faster shortest-path algorithm, C-HD — ctjlewis · 2026-09-23
- Developer claims Anthropic noticed community posts about 'Claudelish' speech quirks — evijit · 2026-09-23
- scaling01 taunts haters after joking Anthropic won't ship Opus 5.5 and Sol today — scaling01 · 2026-09-23
- Delip Rao downgrades from $200/mo Google One Ultra to $50 Pro, leaning on local models — deliprao · 2026-09-23
- Why AI progress accelerated: Claude 4.5 kicked off narrow RSI and open-weight catch-up — maksym_andr · 2026-09-23