Reddit User Tests MiniMax H3 with t2va-then-ref2va Pipeline for Consistent Clips

urabewe · reddit · 2026-10-07

A Reddit user shares a three-clip 10s short test built with MiniMax H3 and Yue2 for music: the first clip uses text-to-video, the rest use ref2va with clips and voice from previous generations for consistency. Post work unified the dog barks, removed generated music and isolated vocals. Settings: 0.6mp, lcm/beta57, 4 steps with the DMAD 4-step LoRA.

Original post →

More from Multimodal

Multimodal channel →