fal launches H3 Max Lip Sync: photo + audio to lip-synced video in 11 seconds
jfischoff · x · 2026-09-19
fal has launched H3 Max Lip Sync: upload one photo and one audio clip and get a lip-synced video in seconds, with natural expressions and movement in any language. It ranks #1 in fal's evals for both quality and speed, with a median generation time of just 11 seconds.
Related event: fal Launches H3 Max Lip Sync, Topping Quality and Speed Benchmarks(4 posts)→
More from Multimodal
- LTX 2.5 CQ Enhancer LoRA Delivers Striking Upscale Results — blastbottles · 2026-09-19
- Where Grok Imagine 2.0 Edits Best: Scene & Style, Identity-Preserving, Restoration — ArtificialAnlys · 2026-09-19
- Grok Imagine 2.0 Strongest at Knowledge, Text Rendering and Reasoning in Text-to-Image — ArtificialAnlys · 2026-09-19
- Grok Imagine Image 2.0 Climbs 14 Spots to #4 on Text-to-Image Leaderboard, Best Non-OpenAI Model — ArtificialAnlys · 2026-09-19
- Modal x Gray Area picks 13 AI art commissions, each with $20K cloud credits — charles_irl · 2026-09-19
- AI Short Film 'Straw Tears' Reveals Full Pipeline: Midjourney, Seedance 2, Magnific — LudovicCreator · 2026-09-19