Choosing a local LLM for coding on RTX 5070 Ti

thatObstinateGuy · reddit · 2026-08-18

A developer with an RTX 5070 Ti (16GB VRAM) and 32GB RAM seeks advice on choosing a local coding LLM. The user is comparing Qwen3.6-35B-A3B and Qwen3.8-27B-UD-IQ3XXS quantized versions, noting that while the latter excels at planning, it struggles with implementation. The user seeks recommendations or methods to benchmark coding performance locally.

Original post →

More from coding & agent

coding & agent channel →