Qwen3.8 27B Runs on Consumer GPUs as Low as 12GB VRAM

Users report running Qwen3.8 27B on consumer hardware: one achieves 18 tok/s decoding with just 12GB VRAM plus 8GB RAM, while another confirms it works on a 16GB 5060Ti using IQ3 quantization with up to 131k context.

2026-10-07 ~ 2026-10-07 · 2 related posts