ComfyUI INT6 Quantization Node Cuts Storage by 25%

BakaPotatoLord · reddit · 2026-08-27

A developer released an experimental ComfyUI custom node implementing INT6 quantization, sitting between INT4 and INT8. This scheme packs 4 bytes into 3 bytes, reducing storage by 25% compared to INT8 with similar generation speeds, as weights are unpacked to INT8 at runtime. The author published a Z-Image-Turbo INT6 model (4.73GB) aimed at saving VRAM/RAM for lower-end GPUs like the GTX 1660 Super while maintaining quality.

Original post →

More from Infra

Infra channel →