QwenImage 2.1 INT4 runs in just 4GB VRAM with ConvRot quantization

reeight · reddit · 2026-09-20

A Reddit user shares low-VRAM QwenImage 2.1 setups: toxicdog's Int8ConvRot quantization with the default ComfyUI T2I workflow and a w4a8 CLIP model. INT8 fits in 8GB VRAM while INT4 runs on just 4GB, making local inference feasible on consumer GPUs.

Related event: QwenImage 2.1 quantized release runs on as little as 4GB VRAM(2 posts)→

Original post →

More from Infra

Infra channel →