Qwen 3.8 27b Local Coding Test: 73 tok/s at 128k Context, Usable Performance

julianharris · x · 2026-08-17

A developer tests Qwen 3.8 27b for local coding. With 4-bit unsloth quant on RTX 4090, it achieves 73 tok/s at 128k context, absolutely usable. Disabling thinking mode gives 78 tok/s but much lower quality. Compared to Qwen 3.6 in May, setup is much simpler. Local AI coding agents look promising.

Original post →

More from coding & agent

coding & agent channel →