User math: peak pricing 2x but base rate better, 252M cache tokens cost just $0.75/day
teortaxesTex · x · 2026-09-09
A heavy user breaks down the new pricing: peak rates are 2x higher, but the base rate is noticeably better than before, so the overall situation improves. What matters for heavy users is essentially only the cache read price, which swamps everything else. The author burned 252M cache tokens in a day, costing just $0.75 at the new prices — under $3 per million tokens, showing how decisive the cache-read discount is for high-frequency agent workloads.
More from Models
- GPT-6 Astra system card draws fire: OpenAI claims 'most aligned model' ever — TheZvi · 2026-09-09
- Muse reportedly makes phone calls in Croatian, but users say it denies the ability — nathanbenaich · 2026-09-09
- DeepSeek V4.1 Flash tops Chinese models in coding blind test, but drops 17 points when switching clients — teortaxesTex · 2026-09-09
- GPT-6 Astra gears up for Singapore F1 night race with live demo — gabrielchua · 2026-09-09
- Nex-N2.5 Goes Free on OpenRouter: 262K-Context Agentic Coding Model with Visual Feedback Loop — airesearch12 · 2026-09-09
- Raschka deep dive: GPT-6 Astra, recurrent depth, and whether it hides its chain of thought — bibryam · 2026-09-09