Qwen 3.8 27B Q3 Review: Strong Performance Despite Low Quantization

AltruisticList6000 · reddit · 2026-08-21

A user reports on running the Qwen 3.8 27B Q3xxs quantization on an RTX 4060 Ti 16GB. Despite typically avoiding Q3 quants, this model impressed by one-shotting serious coding tasks—producing fully working games or web apps—outperforming the previously used Qwen 3.6 35B. It generates at 30-35t/s when fully in VRAM, comparable to the older model's offloaded speed. Downsides include occasional misunderstandings in casual chat or basic counting errors, though math and logic capabilities remain strong.

Original post →

More from Models

Models channel →