Tencent Hunyuan's Prism: dynamic sparse attention speeds up 2K joint video-audio training 2.5x

Tencent-Hunyuan · hf · 2026-10-06

Tencent Hunyuan introduces Prism, a dynamic sparse attention framework for natively training joint video-audio generation models at 2K resolution.

Related event: Tencent Hunyuan Open-Sources Prism for Native 2K Video-Audio Joint Generation(2 posts)→

Original post →

More from Multimodal

Multimodal channel →