Flux1 images to WAN2.2 clips: how to stitch them into 15+ second videos
wreck_of_u · reddit · 2026-08-24
A ComfyUI user describes their character video pipeline: a headless server batch-generates Flux1 images overnight using a reliable character LoRA, followed by manual filtering of body-horror and low-likeness results, then 5-7 second WAN2.1/2.2 clips via Replicate and Vast.
Their core question: can these short clips be concatenated into 15+ second videos with AI seamlessly extrapolating transitions? Should still images or short clips serve as input? Or would training a new LoRA from the images/videos to generate longer videos from text prompts be better? They also ask what tools to use now, mentioning Minimax H3.
More from Multimodal
- Seedance 2.5 launches inside Magnific with stable multi-character video interactions — Div_pradeep · 2026-08-25
- Adobe Firefly expands with music and sound generation — thione · 2026-08-25
- Meituan releases open audio-driven video model — thione · 2026-08-25
- OpenAI adds transparent-background output preview to GPT Image 2 in API — thione · 2026-08-25
- Evangelion's Rei Watches Barney Generated by MiniMax H3 — Certain_Potato_4509 · 2026-08-25
- Magnific Offers Unlimited MiniMax H3 Max for 3 Days, Ultra-Fast Video Generation — cuenca · 2026-08-25