DeepSeek Might Need 100k GPUs to Hit ARR Target
teortaxesTex · x · 2026-07-15
Based on the current pricing and speed of V4, the author estimates that for DeepSeek to reach $8B ARR, it would need roughly 100,000 inference GPUs at peak capacity.
The post also provides rough calculations:
- About 11.5 quadrillion output tokens
- About 230 quadrillion total tokens
- The author believes this scenario is "not impossible"
This post primarily discusses the relationship between inference compute demand and commercial scale, rather than just product opinions.
Related event: DeepSeek inference economics: sizing the GPU fleet behind its ARR(8 posts)→
More from Infra
- Weaviate adds per-query profiling to pinpoint where a slow search query spends time — CShorten30 · 2026-07-22
- Mistral expands its Microsoft partnership as it adds more AI compute in Europe — MistralAI · 2026-07-22
- Sol-Engine Boosts Video Generation Speed by up to 5x with Training-Free Sparse Attention — songhan_mit · 2026-07-22
- PyTorch CTO to Explore Open Source AI Inference Economics and Workflow Optimization — PyTorch · 2026-07-22
- Tinkerers run GLM-5.2 at near-lossless quality on a $15,000 budget — amplifiedamp · 2026-07-21
- AI accelerators now account for 15–20% of active North American data-center power — BenBajarin · 2026-07-21