Choosing a local LLM for coding on RTX 5070 Ti
thatObstinateGuy · reddit · 2026-08-18
A developer with an RTX 5070 Ti (16GB VRAM) and 32GB RAM seeks advice on choosing a local coding LLM. The user is comparing Qwen3.6-35B-A3B and Qwen3.8-27B-UD-IQ3XXS quantized versions, noting that while the latter excels at planning, it struggles with implementation. The user seeks recommendations or methods to benchmark coding performance locally.
More from coding & agent
- Build interactive landing pages in one shot with Gemini 3.7 Flash — jocarrasqueira · 2026-08-18
- Opinion: Bot core split should be IAM, not tasks — willccbb · 2026-08-18
- AI Agent Success Depends on Context Layer, Not Just Models — Pavan_Belagatti · 2026-08-18
- Deploying a Hermes agent on a dedicated phone to automate content tracking — Teknium · 2026-08-18
- Interactive causal network built with Grok 4.6 and vanilla JS: trace, intervene, and reshape systems — techartist_ · 2026-08-18
- Failed AI Agents Consume 40% More Tokens Than Successful Ones — 0xsachi · 2026-08-18