DeepSeek Might Need 100k GPUs to Hit ARR Target
teortaxesTex · x · 2026-07-15
Based on the current pricing and speed of V4, the author estimates that for DeepSeek to reach $8B ARR, it would need roughly 100,000 inference GPUs at peak capacity.
The post also provides rough calculations:
- About 11.5 quadrillion output tokens
- About 230 quadrillion total tokens
- The author believes this scenario is "not impossible"
This post primarily discusses the relationship between inference compute demand and commercial scale, rather than just product opinions.
Related event: DeepSeek inference economics: sizing the GPU fleet behind its ARR(8 posts)→
More from Infra
- OpenRouter agents now out-consume humans as AI usage arrives in three waves — AccBalanced · 2026-09-11
- Nvidia Is Now Core to Every Major Robotaxi Stack at Commercial Scale — pdamodaran · 2026-09-11
- 12 KV Cache Reduction Techniques Every AI Engineer Should Understand, Explained — blaizedsouza · 2026-09-11
- The shadow GPU capacity market is formalizing, with Meta selling excess compute to outside buyers — DavidLinthicum · 2026-09-11
- Engram's random reads don't suit SSDs; CPU-memory over NVLink could serve all 72 GPUs — bookwormengr · 2026-09-11
- RunningHub open-sources H3Lightning, speeding up MiniMax H3 video generation 12x — 智东西 · 2026-09-11