DeepSeek Local Deployment: 4x W7900 vs 2x Blackwell for Agentic Workflows
Retumbo77 · reddit · 2026-08-24
A user is weighing hardware configurations for running DeepSeek 0731 Q4 locally, specifically as an orchestrator for agentic workflows. The options include upgrading to a 6x RTX 3090 rig ($7k), switching to 4x AMD W7900 48GB cards, or purchasing 2x RTX 5000 Blackwell 72GB cards ($14k net). The user seeks insights on expected performance (tokens/sec) for these edge cases, particularly noting that AMD setups might run faster on Vulkan than ROCm.
More from Infra
- Infrastructure shift: Agent CPU scheduling & GPU prefill define new 'Pareto optimal' — AccBalanced · 2026-08-24
- Qwen 2.5 27B Quantization Benchmark: Q4 Viable, Q6 Safest — Fun-Meaning-6474 · 2026-08-24
- LLM Caching Strategies: Reduce Latency, Cost & Model Load — blaizedsouza · 2026-08-24
- Overseas data center shift won't stop AI buildout, costs local economy — QuintinPope5 · 2026-08-24
- GoatRemote stable release integrates Groq for ultra-fast TV control — mayfer · 2026-08-24
- DS4 vs oMLX for DeepSeek V4 Flash on Mac Studio? — ShittyMillennial · 2026-08-24