Max reasoning effort nearly doubles cost for minimal Intelligence Index gains

randal_olson · x · 2026-10-07

Using Artificial Analysis's Oct 6 snapshot, Randal Olson shows max reasoning effort nearly doubles cost per benchmark task for minimal Intelligence Index gains: Claude Opus 5.5 at xhigh ($3.46/task, score 56) roughly matches Sonnet 5.5 at max ($7.67/task, 56) for less than half the cost. Full ladder: Opus 5.5 max $5.98 (58), Sonnet 5.5 max $7.67 (56), Opus xhigh $3.46 (56), Opus high $1.82 (54). Takeaway: compare models before switching to max. Charts made with his open source evident-charts skill.

Related event: Max Reasoning Effort Nearly Doubles Cost for Minimal Gains(2 posts)→

Original post →

More from Models

Models channel →