Which GPU is Best for Local Coding LLMs
Think_Illustrator188 · reddit · 2026-07-09
The author is building a PC for local coding LLMs, prioritizing inference with some lightweight fine-tuning. They are debating between three GPU options: a modded RTX 4090 48GB, two AMD Radeon AI Pro R9700 32GB, or two Intel Arc Pro B70 32GB. The post compares the VRAM, bandwidth, price, drivers, and ecosystem concerns for each.
More from Infra
- oMLX 0.5.2 adds Mac menu-bar stats, low-bit decode kernels, and faster downloads — awnihannun · 2026-07-22
- Strangeworks launches Aura to turn enterprise ops into production optimization systems — whurley · 2026-07-22
- Graph workload 854.graph500 enters SPEC CPU 2026 as a new CPU benchmark — Prof_DavidBader · 2026-07-22
- Hybrid and local inference are emerging as a response to AI energy and token costs — dmitry140 · 2026-07-22
- NVIDIA details Vera CPU with 2x performance claims and a 22,000-core rack — ryanshrout · 2026-07-22
- NVIDIA says Vera Rubin NVL72 delivers 10x more tokens per megawatt than Blackwell — nvidia · 2026-07-22