MiniMax H3 long-video coherence caps out around 20 seconds in community testing
Tight_Organization54 · reddit · 2026-09-05
The author polls the community on the longest coherent single-pass MiniMax H3 video and the setups used (model, turbo, sampler, scheduler, steps). Their best result is 20 seconds with the regular flfva model at 8 steps; pushing to 30–45 seconds causes the video to loop the first couple of shots while adding flavors of later ones. They also ask for long-video workflows beyond Continuum, which rarely works for them.
More from Multimodal
- Two 45-minute loops with a critic: how GPT-6 Astra iterates on 3D reconstruction — bilawalsidhu · 2026-09-05
- GPT-6 Astra rebuilds a living room in Blender from a 3D scan, no downloaded assets — bilawalsidhu · 2026-09-05
- AI VFX test composites into live-action footage, though overhead angles still drift — ssoissoi · 2026-09-05
- Astra first model to produce aesthetically pleasing hand-drawn caricature cutaways — andrew_n_carr · 2026-09-05
- Sketch-to-3D: "GPT-6 Astra" builds a full Blender rocket from a drawing prompt — thursdai_pod · 2026-09-05
- UrbanGround benchmark: LLMs' long-range city navigation drops to 0% — 机器之心 · 2026-09-05