Building a $15k Local Inference Rig for DeepSeek V4 Flash: MI210 vs A40?

_TheWolfOfWalmart_ · reddit · 2026-09-12

A Reddit user seeks GPU advice for a $15k work rig to run DeepSeek V4 Flash and similar models locally with at least 128GB VRAM and minimal quantization, fitting 3 cards in a Dell R740. Candidates: 3× AMD MI210 (192GB) vs 3× NVIDIA A40 (144GB); they're asking for real token-gen and prompt-processing benchmarks and say cloud rental for testing is unavailable.

Original post →

More from Infra

Infra channel →