One repeated sentence keeps a character’s face and voice consistent across shots
Minute_Eye_6270 · reddit · 2026-07-23
The author says they combined JoyAI-Echo’s cross-shot character memory with LTX-2.3’s voice so that one repeated sentence can keep both a face and a voice consistent across every shot.
What’s included
- The workflow used
- Weight formats: bf16, fp8, Q8, Q5, INT8
- A free demo Space
The post is essentially a showcase of a video-generation pipeline that preserves identity across scenes while experimenting with different quantization formats.
Related event: New Workflow Achieves Multi-Shot Audio-Visual Consistency for Characters(2 posts)→
More from Multimodal
- Lumara AI Film Festival Comes to NYC Oct 26, Top AI Filmmakers to Compete — 0xAllen_ · 2026-09-11
- Pterodactyl Detective: An AI-Generated Proof-of-Concept Trailer — PterodactylDetective · 2026-09-11
- Imperium Game Trailer Showcases AI Video Generation — keaslenyt · 2026-09-11
- FLUX.2 Klein Drifts Hard on Character Expressions While Free Gemini Holds Likeness — wacomlover · 2026-09-11
- Tencent Hunyuan releases AuK code and weights on GitHub with ComfyUI and fine-tuning support — aigclink · 2026-09-11
- Creator turns Bahamut vs Tiamat rivalry into an AI cinematic battle with Midjourney, GPT Image 2 and Seedance — azed_ai · 2026-09-11