FLUX.1-dev ConvRot conversion cuts peak VRAM by up to 32.5%
ThaJedi · reddit · 2026-07-22
FLUX.1-dev ConvRot ports cut peak VRAM by up to 32.5%
The author converted FLUX.1-dev into native ComfyUI ConvRot formats and shared VRAM measurements at 1024² / 20 steps.
Key results:
- Partial INT8: 24.09 → 20.35 GiB, a 15.5% reduction
- Whole W8A8: 16.30 GiB, a 32.3% reduction
- W8A8 + INT8 T5: 16.27 GiB, a 32.5% reduction
The model is posted on Hugging Face and Civitai, with the main value being lower peak memory for high-fidelity runs.
More from Infra
- AMD and Anthropic reportedly strike a 2-gigawatt chip deal worth tens of billions — kimmonismus · 2026-07-22
- AMD and Anthropic sign a multibillion-dollar chip deal for 2 gigawatts of compute — kimmonismus · 2026-07-22
- SRAM costs are closing in on HBM, and 3D SRAM accelerators may finally have a market — zephyr_z9 · 2026-07-22
- How to install PyTorch with uv across CPU, CUDA, ROCm and XPU — tdhopper · 2026-07-22
- ComfyUI Wan2.2 i2v workflow may no longer keep models cached in RAM — Fun-Class3451 · 2026-07-22
- AMD Strix Halo user builds llama.cpp Laguna on ROCm and hits flash-attn limits — dbinnunE3 · 2026-07-22