Estimating Compute Revenue and GPU Scale
teortaxesTex · x · 2026-07-15
Based on the "bundle cost," the author estimates that peak load requires about 1000–2000 seconds of GPU time.
Calculated against this metric, the revenue per GPU is roughly $28,000 to $55,700 annually, corresponding to approximately 9,000–18,000 GPUs, or 1–2 SuperPODs. The author adds that if peak loads are more extreme and the 2K estimate is too low, it might require up to 40,000 seconds of GPU time, suggesting the actual scale could be even larger.
Related event: DeepSeek inference economics: sizing the GPU fleet behind its ARR(8 posts)→
More from Infra
- Ratel says it made agents 7x cheaper by loading only the tools each task needs — tensorqt · 2026-07-22
- Weaviate adds per-query profiling to pinpoint where a slow search query spends time — CShorten30 · 2026-07-22
- Mistral expands its Microsoft partnership as it adds more AI compute in Europe — MistralAI · 2026-07-22
- Sol-Engine Boosts Video Generation Speed by up to 5x with Training-Free Sparse Attention — songhan_mit · 2026-07-22
- PyTorch CTO to Explore Open Source AI Inference Economics and Workflow Optimization — PyTorch · 2026-07-22
- Tinkerers run GLM-5.2 at near-lossless quality on a $15,000 budget — amplifiedamp · 2026-07-21