Open-source LongCat-Avatar turns one photo plus audio into minutes-long lip-synced talking video
anthara_ai · x · 2026-10-08
A viral post highlights LongCat-Avatar, an open-source model that generates minutes-long lip-synced talking-head video from a single photo plus an audio clip — capability previously gated behind paid AI video tools.
The author argues it collapses what once required a camera, studio, and editing into a single free repo, and links to the project from Meituan's LongCat team.
More from Multimodal
- LoCoSplat ditches learned 3D networks: 4.2x faster, 6.7x less memory feed-forward 3DGS — zhenjun_zhao · 2026-10-08
- Nano Banana 2.1 wows users in latest community image generation tests — miilesus · 2026-10-08
- AI-Generated Short Film 'BOTCHED: The Christa Pike Story' Made with Seedance 2.5 — AIandDesign · 2026-10-08
- Redditor Releases Part 2 of Dark Fantasy AI Series 'THE FOREST' — FauxFeast · 2026-10-08
- Midjourney posts Office Hours recording from October 7 — midjourney · 2026-10-08
- Local 3B music model writes the track, Claude Code makes the MV in 15 minutes from one prompt — Pleasant_Salt6810 · 2026-10-08