DeepSeek reportedly runs smaller-model inference on NVIDIA gaming GPUs, argues Teortaxes

teortaxesTex · x · 2026-09-26

Teortaxes cites a claim that DeepSeek is using NVIDIA gaming GPUs to run inference on smaller models, calling it an obvious move that pairs well with Jevons-style small models or a custom Qwen 35B. The quoted post from tugot17 argues that as RSI nears, labs' demand for frontier chips has no upper cap and everyone else gets priced out — making heterogeneous compute unavoidable.

Related event: Report: DeepSeek Runs Small-Model Inference on RTX 5090s, Draining Retail Supply(2 posts)→

Original post →

More from Infra

Infra channel →