Local AI Pipeline Demo: IndexTTS 2.5 DoRA Voice + MiniMax H3 Image + Lip-Synced Video

CeFurkan · reddit · 2026-10-02

Reddit user CeFurkan shared their highest-quality local AI video yet, combining a DoRA fine-tuned IndexTTS 2.5 voice clone, MiniMax H3 image generation, and speech-to-lip-sync. The full demo is on YouTube, with the author soliciting feedback on voice and avatar quality.

Original post →

More from Multimodal

Multimodal channel →