MiniMax H3 long-video coherence caps out around 20 seconds in community testing

Tight_Organization54 · reddit · 2026-09-05

The author polls the community on the longest coherent single-pass MiniMax H3 video and the setups used (model, turbo, sampler, scheduler, steps). Their best result is 20 seconds with the regular flfva model at 8 steps; pushing to 30–45 seconds causes the video to loop the first couple of shots while adding flavors of later ones. They also ask for long-video workflows beyond Continuum, which rarely works for them.

Original post →

More from Multimodal

Multimodal channel →