Dual RTX 5080 and 128GB RAM: What Models to Run on 32GB VRAM
whatyathinkk · reddit · 2026-09-07
A user set up a consumer tower with dual RTX 5080s (32GB VRAM total) and 128GB DDR5 RAM on a Ryzen 7 7700X, asking which models people run on similar setups and hoping for exact llama.cpp settings. They're also considering upgrading to an Epyc CPU for memory bandwidth.
More from Infra
- Microsoft open-sources tgrep, a trigram-indexed grep up to 52x faster than ripgrep — jedisct1 · 2026-09-07
- How should billing work when an AI system auto-selects the model? — Colddew-YJ · 2026-09-07
- Can you run Qwen Next on a 3090 + 64GB CMP 170HX? Local deployment help — JustinPooDough · 2026-09-07
- SmolVM: open-source microVM sandbox runs OpenClaw 2.0 in isolation, boots in milliseconds — aniketmaurya · 2026-09-07
- Hesamation recommends the best technical book on training LLMs at scale — free to read — Hesamation · 2026-09-07
- After His OpenAI Key Was Stolen, He Found Stratum: a Docker-Layer Secret Scanner Crunching 700K Layers Daily — Ubunta · 2026-09-07