Seamless action scenes in ComfyUI: FL2V + Minimax H3 guide workflow on a 3090

VisionWithin · reddit · 2026-08-27

The author experiments with ComfyUI's FL2V node plus "Add Guide for Minimax H3" for seamless action scenes: a one-second guide video is enough for a visually seamless result, though audio must be produced in post, guided only by the video frames and prompting.

Everything was made with 16:9 480p videos and an 8-step turbo LoRA; one generation takes 2.5 minutes on an RTX 3090, leaving plenty of time to redo shots and refine prompts. Main limitation: no subject consistency except by chance—more serious work needs the Ref2V node for guided generation. Verdict: a lot of fun.

Original post →

More from Multimodal

Multimodal channel →