ComfyUI video workflow shared: character refs + voice sync for consistent shots
boricuapab · reddit · 2026-09-06
A creator shared the full ComfyUI workflow behind one of their sync challenge shots: the h3 t2va model combined with the ref2v acc LoRA and an LTX2.5 upscaler, using two character reference images to keep the character consistent across shots.
For audio consistency, they recorded their own voice as the audio reference for the male character, aligning lip movement and voice. The workflow file (including the two character refs) is downloadable via Ko-fi — you only need to record your own voice to reproduce it.
More from Multimodal
- Yoroll launches YoLive, a never-ending AI livestream film with 4s video generation — vista8 · 2026-09-06
- Google brings Lyria 3.5 AI music generation directly into the Gemini app — The Decoder · 2026-09-06
- Running Qwen 3.8 27B Locally to Drive Blender via MCP and Build a 3D Llama — jacek2023 · 2026-09-06
- AI-Generated Jeans Ad Video Shows Near-Photoreal Commercial Quality — Foreign-Original124 · 2026-09-06
- Meta's Muse Voice Transcribe: 80ms streaming transcription, cheapest and most accurate — The Decoder · 2026-09-06
- Electronic musicians return to hardware as AI threatens to replace the creative process — teropa · 2026-09-06