Is Qwen 3.8 27B at Q2 quantization still usable? A 16GB owner asks
Effective_Head_5020 · reddit · 2026-08-29
A Reddit user runs Qwen 3.8 27B at Q3 on an AMD RX 9060 16GB and it works, but heavy thinking blows up the context, so they're considering Q2 or Q3 IQ xxs. They cite Luke's dev lab video showing Q2 is decent, and ask for real-world impressions.
Related event: Quantized Qwen 3.8 27B Runs 200K Context on 16GB VRAM(3 posts)→
More from Models
- User Reports X Platform May Have Removed Grok Auto-Reply Feature — BLUECOW009 · 2026-08-29
- GLM 5.3 Flash Achieves Lowest Hallucination Rate — SumitGup · 2026-08-29
- Kimi-K3 prices rise 30% due to high demand — markjeffrey · 2026-08-29
- GLM-5.3 Launches on Baseten: 743B Open-Weight Model with ZDR — baseten · 2026-08-29
- GLM-5.3 launches on Baseten with 743B params and MIT license — baseten · 2026-08-29
- MoE Inference Analysis: Qwen 27B vs. Flash-Next 288B on M5 Pro — EyalToledano · 2026-08-29