Minimax H3 tip: blend latent noise to loosen overly literal reference images
neph1010 · reddit · 2026-09-16
A ComfyUI user shares a hands-on technique for Minimax H3's overly literal ref2va image references: instead of feeding the raw image, blend its latent with noise. Noise in latent space gives the sampler freedom to explore nearby concepts. Empirically, latent strength 0.35 is the sweet spot (keeps composition, regenerates texture); 0.38 stays blocky, 0.25 drifts toward no-reference output, and the drop-off is sharp. Since native nodes don't accept latent input, the author released custom modified nodes on GitHub; the same trick works well with Flux2 Klein's latent-as-reference.
More from Multimodal
- Seedance 2 video demo shows off multi-talented AI-generated performer — Ok-Nerve941 · 2026-09-16
- Dubstep music video created with LTX 2.3 and Suno — OkDifference4231 · 2026-09-16
- Redditor impressed by AI-animated video, imagines a decade ahead — 8th_circle · 2026-09-16
- Boreal by Creatify lands on fal: prompt-to-ad videos at 2K, up to 20s — Scobleizer · 2026-09-16
- YuE2 on an RTX 5070 Ti generates a 4-minute song in 130 seconds — Downtown-Cover-7422 · 2026-09-16
- LAION-backed voice acting arena lets you blind-compare TTS models performing the same scene — realmrfakename · 2026-09-16