Benchmarking DeepSeek on Dual RTX 3060 with 96GB RAM: Speed and Cost Analysis
esw123 · reddit · 2026-08-02
A developer shared benchmark results running a DeepSeek model using dual RTX 3060 GPUs and 96GB of RAM via Unsloth Studio.
- Hardware: Ryzen 7500F CPU, two RTX 3060s (on PCIe 5.0 x16 and PCIe 3.0 x1 lanes respectively), and 96GB 5600MHz RAM.
- Performance: The generation speed hovered around 3.5 tok/s, taking roughly 16 minutes to generate 4,338 tokens.
- Power & Cost: Both GPUs were undervolted to 0.9V (30-40W each). The total task consumed an estimated 35 watts, costing approximately €0.0059 in electricity.
More from Infra
- AMD MI355X Beats NVIDIA B200 in Kimi K3 Deployment with 952 tok/s — adrianscottcom · 2026-08-03
- MiniMax H3 Gets Day 0 Support in SGLang, Runs Locally on Dual RTX 5090s — ying11231 · 2026-08-03
- Qwen3.8-27B Open Weights Coming, Runs Locally on 17GB RAM — danielhanchen · 2026-08-03
- AirLLM Breaks VRAM Barrier: Runs 70B LLMs on a Single 4GB GPU — techNmak · 2026-08-03
- MiniMax H3 Open Weights Hit fal with Out-of-the-Box Inference Optimizations — gorkem · 2026-08-03
- ComfyUI Adds Day 0 Support for MiniMax Video Model, Slashing VRAM by 66% for RTX 3060 — crystal_alpine · 2026-08-03