Anthropic's own cost guide includes a lever that cost 74% more
Brinvik · reddit · 2026-09-10
A lever in Anthropic's newly published cost guide — context editing and compaction — saved nothing and cost 74% more on a 20-issue run, yet saved 39% and 32% on longer runs. The sign flips with run length, meaning the guide is a set of bets on your workload rather than guaranteed wins, and someone must measure before changing anything.
Other figures from Anthropic's own benchmarks:
- Prompts written for Opus 4.8 cost 36% more per support ticket on Opus 5 with no accuracy gain; removing "verify twice" alone cut cost per ticket by a third on Opus 5.
- Agent loops read a median 84% of input from cache over a day of real traffic; top 10% of setups hit 94%+.
- Caching cut agent-loop cost by 2.7x to 5.3x.
- Without pauses, the five-minute cache default cost 15% less than the one-hour setting on Sonnet 5 and 11% less on Opus 5 — longer caches were pricier.
The author flags mismatched baselines (36% is Opus 4.8 vs Opus 5 untouched prompts; 14% is audited vs unaudited on Opus 5) and notes accuracy-restoring fixes (a retired thinking setting, contradictory rules, a hand-rolled scratchpad) each gave back 7-11 points — bigger than most cost levers.
More from coding & agent
- Automated UI testing in Devin 'feels like AGI': agent plans, clicks, records and edits test videos autonomously — smtabatabaie · 2026-09-10
- Qwen Code Desktop v0.3.0 ships session restore and Windows PTY fixes — github-actions[bot] · 2026-09-10
- Hermes Agent now supports installing plugins from private GitHub repos — Teknium · 2026-09-10
- Salesforce paper: fine-tuning on evolved harnesses makes weak models worse on all 7 tasks — omarsar0 · 2026-09-10
- "I wanted a bicycle for the mind, we got DoorDash for Thinking" — round · 2026-09-10
- 20-Year Veteran Editor: GPT-6 Astra Edits in DaVinci via MCP, "Better Than Me" — petewoodbridge · 2026-09-10