AI token demand could reach 120 quadrillion a month by 2030, post argues
bittingthembits · x · 2026-07-29
Goldman Sachs is being used here as a backdrop for an argument that AI token demand could explode by 2030.
- The post cites a projection of 120 quadrillion AI tokens per month by 2030, then argues the real number could be several times higher if agents keep running continuously.
- It compares current token prices across frontier models and smaller inference providers, then extrapolates what that usage would mean at different price points.
- At $1 per million tokens, the author estimates $120B/month of inference spend.
- At an assumed average of $0.10 per million tokens, the estimate becomes $360M–$480M/month, or $4.3B–$5.8B/year.
- The core thesis is that agent workloads won’t be one-off requests; they’ll keep calling models, so token volume and inference revenue could be much larger than today’s chat usage.
More from Infra
- Gavin Baker says compute rental prices may keep rising as AI demand stays hot — GavinSBaker · 2026-07-30
- LiveKit says Gemma 4 31B hits 192 ms to first token in voice agents — GlennCameronjr · 2026-07-30
- Optimized Qwen Image 2512: 5x Smaller, 3x Faster Inference — enrique-byteshape · 2026-07-30
- Reddit user gets about 4 tokens/s running Kimi K3 on a 2×5090 home lab — iVoider · 2026-07-30
- AI Infrastructure Stocks Cool Down: CRWV at 52-Week Low, NVDA Down 8.69% — GaryMarcus · 2026-07-30
- Two RTX 3090s still struggle to fit Qwen Image Edit alongside a 27B text model — Civil_Fee_7862 · 2026-07-30