Tencent Hunyuan open-sources Prism: sparse attention for native 2K joint video-audio generation
Diabolicor · reddit · 2026-10-06
Tencent's Hunyuan team has released code for Prism (Dynamic Sparse Attention for Native 2K Joint Video-Audio Generation), a training approach for native 2K joint video and audio generation using dynamic sparse attention to cut training cost.
- HuggingFace: FrancisRing/Prism
- GitHub: Tencent-Hunyuan/Prism
- A demo video is available on YouTube
Related event: Tencent Hunyuan Open-Sources Prism for Native 2K Video-Audio Generation(3 posts)→
More from Multimodal
- inclusionAI open-sources Ming-Image-0.1-Design: 6B text-to-image model with transparent RGBA output — lmoroney · 2026-10-06
- Local Minimax H3 2K video generation on a 5090: 1984x1120 in 30 minutes, full prompt shared — Kooky-Mode3047 · 2026-10-06
- Wan2GP memory management update: 15s video now renders in 5 minutes, first-click OOM fixed — orangpelupa · 2026-10-06
- MIRRORSIDE: a behind-the-scenes film set that never existed, made with AI — StrategyMedium5907 · 2026-10-06
- Qwen-Image 2.1 edits look unfinished: user shares ComfyUI params seeking fixes — Suspicious_Aide2697 · 2026-10-06
- Creator hands repetitive workflow to Codex and GPT-6 Astra, keeps creative calls — socialwithaayan · 2026-10-06