FLUX.1-dev ConvRot conversion cuts peak VRAM by up to 32.5%
ThaJedi · reddit · 2026-07-22
FLUX.1-dev ConvRot ports cut peak VRAM by up to 32.5%
The author converted FLUX.1-dev into native ComfyUI ConvRot formats and shared VRAM measurements at 1024² / 20 steps.
Key results:
- Partial INT8: 24.09 → 20.35 GiB, a 15.5% reduction
- Whole W8A8: 16.30 GiB, a 32.3% reduction
- W8A8 + INT8 T5: 16.27 GiB, a 32.5% reduction
The model is posted on Hugging Face and Civitai, with the main value being lower peak memory for high-fidelity runs.
More from Infra
- Spomin: live KV cache compaction squeezes 500k tokens of context into 180k resident — wgaca2 · 2026-09-11
- PiPNN nearest-neighbor search wins three awards, up to 78x faster index building — khademinori · 2026-09-11
- M.2-Oculink eGPU Link Silently Downgrades to PCIe Gen1 — Here's How to Check — El_90 · 2026-09-11
- DeepSeek launches V4.1-Flash with 1M-token context and 4x smaller KV-cache — matlabulous · 2026-09-11
- What Can You Still Run on 8GB VRAM? User Asks for Small Models With Tool Use — riceinmybelly · 2026-09-11
- Spain's hourly 80% renewable matching rules clash as France fast-tracks 700MW sites, UK cuts grid queues — eherrerosj · 2026-09-11