ComfyUI Supports Fast INT4 Inference

Limp-Chemical4707 · reddit · 2026-07-10

The post introduces a new custom node pack for ComfyUI, **ComfyUI-INT4-Fast**, designed to natively run ultra-fast, VRAM-friendly INT4/W4A4 models. The author shared real-world tests: running a pre-quantized Krea2 Turbo model on an RTX 3060 6GB took about 17.64 seconds for a single 1024×1024, 8-step generation. They noted that the node supports mixed-precision checkpoints, automatically routing sensitive layers to the most optimal execution path.

Related event: ComfyUI Introduces INT4 Plugin for Low-VRAM Fast Inference(2 posts)→

Original post →

More from Infra

Infra channel →