Seeking advice for local dev setup on dual RTX 6000s

alexp702 · reddit · 2026-08-31

A user with a Threadripper (128GB RAM) and two RTX 6000 Pro Max-Q GPUs wants to set up a local coding harness for a small team. They are considering GLM, Deepseek R4, and Qwen next, preferring models with vision capabilities. Since quantization is needed for dual-card deployment, they are asking for advice on model choice and VLLM configurations.

Original post →

More from coding & agent

coding & agent channel →