NVIDIA Releases Audio2Face-3D for Audio-Driven Facial Animation
tom_doerr · x · 2026-08-23
NVIDIA has released Audio2Face-3D, technology that generates high-fidelity 3D facial animation from an audio source, supporting both pre-recorded files and real-time streams.
Core Capabilities:
- Accurate Lip-sync: Synthesizes precise motion of the jaw, tongue, and eyes based on phonetic analysis.
- Nuanced Emotion: Infers subtle emotional expression from speech tone, combined with detailed facial skin deformations.
- Flexible Driving: Drives character performance via direct mesh deformations, joint transformations, or blend shape weights.
Toolkit:
The hub provides pre-trained models, a development SDK, plugins for Autodesk Maya and Unreal Engine 5, and sample datasets.
More from Multimodal
- Flashback: exploring immersive virtual worlds in latent space via prompts — nptacek · 2026-08-23
- H3 Video Generation Test: Why Does Character Control in Complex Scenes Rely on Luck? — Hdfjds · 2026-08-23
- Minimax Leads in Prompt Adherence, Flux 3 Wins on Atmosphere and Texture — Grinderius · 2026-08-23
- One Person, Three Weeks: AI-Generated Sci-Fi Film SPECTRUM Released — HashemGhaili · 2026-08-23
- Short film created using Minimax H3 audio and ComfyUI r2v workflow — Repulsive-Rush3505 · 2026-08-23
- Sora-level VFX Spellcasting: Emissive Runes and Zero-G Physics — azed_ai · 2026-08-23