Qwen Releases UniSwap: Streaming Audio-Visual Identity Swapping Model
QwenBusinessUnit · hf · 2026-08-14
Alibaba's Qwen team released the UniSwap model, designed for talking videos to achieve synchronized appearance and voice replacement.
UniSwap employs a unified streaming audio-visual diffusion transformer architecture, coupled with specialized training and inference adaptations to maintain high-quality audio-visual sync and identity transformation during video generation.
More from Multimodal
- Generating 4K Gym Vlog with Seedance 2.5: Full Prompt & Upscaling Workflow — SimplyAnnisa · 2026-08-14
- Midjourney Prompt: Horseback Visitor via Fisheye Doorbell Cam — tisch_eins · 2026-08-14
- Elon Musk Showcases Grok Imagine's Stunning Auto-Animation Feature — elonmusk · 2026-08-14
- Qwen Releases LiveAnimate: 14B Real-Time Streaming Human Animation Model — QwenBusinessUnit · 2026-08-14
- Running MiniMax Video Model on RTX 5090 Uses Only 20GB VRAM — BoredHobbes · 2026-08-14
- Midjourney version 8.2 released — azed_ai · 2026-08-14