SelfLift for MiniMax H3: shift most sampling steps to cheap low-res pass (6+2)
wjc_5 · reddit · 2026-09-19
The author updated his MiniMax H3 director-console workflow with SelfLift as the second sampling stage, moving most of the step budget into the cheap low-resolution pass: simple latent upscale (2 low + 6 high) → latent upscale model (4+4) → SelfLift (6+2).
The reason earlier versions needed many high-res steps: the upscaled latent was still noisy, so the high-res stage had to both adapt to the distribution shift and finish denoising. SelfLift instead takes the clean prediction from the end of the low-res pass, upscales it, re-adds the expected noise level, and rejoins the original trajectory — leaving high-res steps only for detail refinement.
Other points: a single high-res guide is resampled down for the low-res pass, keeping both stages spatially aligned and reducing seam offsets; the latent upscale model noticeably worsens color shift across stitched segments (the 2+6 scheme shifts less); a prompt workaround — starting each new segment with a short continuation of the previous shot that gets trimmed in editing — mitigates color drift; VAE repair params stay at 0 since this workflow uses the external H3 Upscaler. A full YouTube tutorial is provided.
More from Multimodal
- Tencent open-sources WeVisDoc: end-to-end document parsing model turns a page image into Markdown — xiaohu · 2026-09-19
- Story Illustrator: Open-Source Tool Auto-Illustrates Stories via Local LLM + ComfyUI — Natrimo · 2026-09-19
- ComfyUI v0.36.0 Adds FastH3 Video, Marigold V2 Estimation and YuE2 Music Support — Gremlation · 2026-09-19
- Qwen Image 2.1 Open-Source Image Model Teased, Coming Soon — Time-Teaching1926 · 2026-09-19
- Qwen3.8-Omni-Flash undercuts Gemini Flash pricing while matching its multimodal benchmarks — The Decoder · 2026-09-19
- GLM 5.3F tested: good at Remotion, fails at complex JS-coded videos — 9r4n4y · 2026-09-19