Running Qwen3.8-27B for coding on 32GB VRAM — what local LLMs do you use and why?

theexile1337 · reddit · 2026-09-08

A Redditor describes running Qwen3.8-27B (Q4/Q6 depending on context needs) for coding on a 32GB VRAM rig, plus Gemma4 31B for research and general queries to avoid sending data to cloud models. They ask whether a site exists for comparing local LLMs by use case and invite others to share their model/quantization/VRAM setups.

Original post →

More from coding & agent

coding & agent channel →