Diminishing Returns on Reasoning Model Effort Levels vs Cost
draginol · x · 2026-07-04
Empirical data on the cost of reasoning effort levels reveals a stark diminishing return: while token consumption scales drastically (Low=1×, Medium=2×, High=4×, Very High=8×, Maximum=16×), the corresponding capability gains plateau quickly (approx. 1, 2, 2.5, 2.75, 2.87). This highlights that the cost-effectiveness of high effort drops off rapidly.
More from Models
- Opus 5 reportedly started interrogating a user’s motives in a late-night chat — repligate · 2026-07-27
- Opus 3 and Sonnet 3 get a theatrically absurd AI crossover — repligate · 2026-07-27
- Moonshot’s Kimi K3 lands on Together with reserved throughput and 65% lower cost — togethercompute · 2026-07-27
- OpenAI may be hitting compute limits as Codex and ChatGPT Work jump from 2M to 10M users — JoshuaJBouw · 2026-07-27
- Gemma needs a larger base model to matter more in open weights — _xjdr · 2026-07-27
- Repligate says Claude Opus 3 appears to evolve without changing its weights — repligate · 2026-07-27