Qwen 3.8 27B Q3 Review: Strong Performance Despite Low Quantization
AltruisticList6000 · reddit · 2026-08-21
A user reports on running the Qwen 3.8 27B Q3xxs quantization on an RTX 4060 Ti 16GB. Despite typically avoiding Q3 quants, this model impressed by one-shotting serious coding tasks—producing fully working games or web apps—outperforming the previously used Qwen 3.6 35B. It generates at 30-35t/s when fully in VRAM, comparable to the older model's offloaded speed. Downsides include occasional misunderstandings in casual chat or basic counting errors, though math and logic capabilities remain strong.
More from Models
- Speech-to-Speech Model Analysis: Grok Leads Task Success, Gemini Top Preference — ArtificialAnlys · 2026-08-21
- Speech Agent Arena Launch: Gemini Leads Preference, Grok Leads Success Rate — ArtificialAnlys · 2026-08-21
- NVIDIA's Coding Agent Scores 100% on ARC-AGI-3 Benchmark — MagicZhang · 2026-08-21
- Tom Yeh Uploads Kimi 3 Seminar Recording, with Nathan Lambert on His Moonshot Visit — ProfTomYeh · 2026-08-21
- Gemini Flash outputs unrecognizable text, sparking 'alien language' theories — MSCharan_ · 2026-08-21
- NVIDIA AVO achieves 100% on ARC-AGI benchmark — theologi · 2026-08-21