Best local coding model for 16GB RAM? Redditor seeks real-world picks over benchmarks
ImBadGuyInEveryStory · reddit · 2026-09-12
A Reddit user asks for local coding model recommendations under a 16GB system RAM limit, finding Qwen3 27B great but unrunnable on their hardware. They want good code quality, reliable terminal/tool use and agent tool calling, few hallucinations, no infinite loops, and snappy inference—and are especially curious about the best 4B-8B models, favoring real-world experience over benchmarks, plus quantization/runtime choices.
More from Models
- Sentry Founder: Newer 'Smarter' Models Are Producing Worse Results — zeeg · 2026-09-12
- Dev: Astra is a downgrade from Sol unless you brute-force with massive parallelism — mertdumenci · 2026-09-12
- Gary Marcus slams OpenAI for weakening safety monitoring as White House stays silent — GaryMarcus · 2026-09-12
- Dev Fine-Tunes Qwen3.8-27B on 125K Real Conversations to Kill the AI Assistant Vibe — kvyb · 2026-09-12
- Muse, Instinct and Grok Bot 'would be best products of the year' in any other year — jeff_weinstein · 2026-09-12
- Microsoft launches MAI-Transcribe-2: single multilingual transcription model with diarisation and timestamps — mustafasuleyman · 2026-09-12