Minimax H3 in Practice: Ref2VA Workflow Generates 10s Clips in 2-4 Minutes
FewApple7788 · reddit · 2026-09-03
A hands-on workflow for Minimax H3 video generation: Ref2VA with 2-4 reference images per run, scale set to 0.7, outputting 1152×640. On a 5090 with 64GB RAM, each 10-second clip takes 2-4 minutes depending on image and audio refs (using saga). The author is happy with the results.
More from Multimodal
- Claude writes three.js code, Higgsfield renders it: text-to-3D architecture pipeline — heypearlai · 2026-09-03
- Seedance 2.5 turns a supermarket complaint into a AAA game boss fight, full prompt shared — azed_ai · 2026-09-03
- Seedance 2.5 demo turns a retail complaint into a AAA game boss fight — azed_ai · 2026-09-03
- How Squad made its launch video with Revid CLI and 98 script revisions — tibo_maker · 2026-09-03
- Snap a photo, drop it into Blender 3D: the two-step photo-to-3D trick — sidahuj · 2026-09-03
- Topaz Brings Video Enhancement to the Browser: Upscaling, Frame Interpolation, SDR-to-HDR Without Any App — umesh_ai · 2026-09-03