Token Prices Fell 99% in 2.5 Years Yet Total Spend Rose; H100 Spot Climbed All Along
le_james94 · x · 2026-09-26
Lecture 2 of Stanford's MS&E 435:
- Cost per token fell 99% over 2.5 years, yet total spend went up.
- Reasoning models and agents burn orders of magnitude more tokens per task, and H100 spot prices climbed through the entire decline.
- Takeaway: a fast chip is not a business — demand inflation eats the price decline.
More from Infra
- 1Cat-vLLM fork revives decade-old V100s: Qwen3.6-35B benchmark numbers shared — Miserable-Dare5090 · 2026-09-26
- DeepSeek reportedly runs smaller-model inference on NVIDIA gaming GPUs, argues Teortaxes — teortaxesTex · 2026-09-26
- Price of intelligence collapsing 13x per year, says blogger citing Epoch AI data — LeviTurk · 2026-09-26
- Moody's Warns Anthropic and OpenAI Carry $2.5T in Off-Balance-Sheet Debt — SumitGup · 2026-09-26
- SemiAnalysis Maps 1,000+ China Datacenters Across 60+ Operators in New AI Infrastructure Model — zephyr_z9 · 2026-09-26
- Everything About Hosting Got Cheaper Except Moderation: LessWrong Burns ~$1M/Year — jd_pressman · 2026-09-26