ComfyUI video generation experiments: image and audio reference workflow
TensorTinkererTom · reddit · 2026-08-26
The author shares MM (reference image and audio) experiments created with ComfyUI. Despite rough editing and frequent audio glitches, the post details the workflow: primarily a 20-step out-of-the-box ComfyUI workflow with an audio node added. The Load Video node was used to grab the last 24 frames of the previous video to continue the sequence, with limited success. The author notes that a 4-step LoRA was introduced halfway through; even with low weight, audio glitches remained frequent, though video quality was generally decent. Links to sample clips and character swap prompts are provided via Pastebin.
More from Multimodal
- Luma-generated short film "L'heure Bleue" showcases cinematic lighting — mrjonfinger · 2026-08-26
- AI-generated animated short film "Creatures of Habit" showcased — EggRepresentative199 · 2026-08-26
- LTX 2.5 test video shared — call-lee-free · 2026-08-26
- Fun images generated by GPT-Image-2 model — sloppenheimer · 2026-08-26
- What details instantly make you recognize an AI-generated image? — NextHeat8167 · 2026-08-26
- "Better Avoid Saul 3" created with Minimax H3 Image-to-Video — JamesFilmsYT · 2026-08-26