Qwen 27B 3.8 low quantization tested: Q3 XXS works well locally
jeremyckahn · reddit · 2026-08-24
The author shares test results of highly quantized Qwen 27B 3.8 models on limited hardware (Mac mini M4 24GB).
- Performance: Using Unsloth's Q3 XXS quant, the model makes minor mistakes but can work autonomously towards a goal for hours.
- Context: Context window size is around 180k.
- Verdict: For users with lower-end GPUs, this level of quantization is reliable for completing tasks. The author asks if others have had success with sub-Q3 quants.
More from Models
- Claude's invisible watermarks cracked within hours; override code gets 20k bookmarks — deliprao · 2026-08-24
- Sonnet 4.5 exhibits intense, strange behavior in response to Opus 3 — repligate · 2026-08-24
- A comprehensive ranking of various AI models has been shared — FinanceYF5 · 2026-08-24
- User comparison finds LTX outperforms H3 in instrument generation energy — cocktailpeanut · 2026-08-24
- Controversial AI Model Ranking: Fable 5 at S+, Kimi K3 and DeepSeek V4 Flash in Tier B — FinanceYF5 · 2026-08-24
- Video Gen Consumes 70% of AI Tokens in China, Diverging from US LLM Focus — AccBalanced · 2026-08-24