MiniMax H3 Dual-Sampling Workflow: Latent Upscaling & Audio Restoration Guide
wjc_5 · reddit · 2026-08-23
The author tested an updated MiniMax H3 workflow to address issues like limited upscale factors, waxy live-action results, weak audio, and high VRAM usage.
- Latent Upscaling Model: Replacing basic latent resize with a model-based latent upscale eliminates line-like artifacts and fragmented shapes after the second pass. A 1.0x alignment step is recommended for non-integer factors.
- Dual-Sampling Setup: Using the 8-step LoRA for the first pass and 4-step LoRA for the second pass preserves high-motion clothing edges better and avoids character duplication compared to using the 4-step LoRA too early.
- Visuals & Audio: Lowering LoRA strengths (0.75 for 8-step, 0.7 for 4-step) reduces the waxy look in live-action videos. Adding a separate voice-reference audio clip improves speech quality.
- Sampling: Euler with the Beta scheduler works well.
Note that while dual sampling speeds up the first part, it does not reduce peak hardware requirements. A detailed tutorial is available on YouTube.
More from Multimodal
- LTX 2.5 Generation Demo: Using First and Last Frames — waterarttrkgl · 2026-08-23
- Bringer of Shadows: A New Dark Romance Universe — Specialist_Address28 · 2026-08-23
- Claude beats ChatGPT in Fintech UI design prompt test — Tegadesigns · 2026-08-23
- Indie dev let Claude compose the music for their arcade side project — Gambo7592 · 2026-08-23
- Physics explainer video on photons-to-image made entirely with Grok on a phone — yunta_tsai · 2026-08-23
- Ox Alpha recreates 3D glass webpage from a single image using WebGL — op7418 · 2026-08-23