Match Claude model and effort level to the task to cut token usage 2-3x

ifioknkem · x · 2026-10-02

Most people run every task on one model with High or Max effort, burning limits 2-3x faster than needed. The author's setup: Haiku 4.5 for simple tasks and formatting, Sonnet 5 for everyday work and writing, Opus 5 for complex reasoning and strategy, Fable 5.1 only for the hardest jobs. Effort levels: Low for quick summaries, Medium for daily tasks, High for complex analysis, Max only as a last resort. The core rule: deliberately match both model and effort to task difficulty.

Related event: Three Tricks to Cut Claude Token Usage in Half(2 posts)→

Original post →

More from Models

Models channel →