SANA-Video 2.0 pairs hybrid attention with 84.30 VBench and major speedups
danijarh · x · 2026-07-24
SANA-Video 2.0 is a newly released video model optimized end-to-end for efficiency while keeping quality high.
It uses a hybrid linear-softmax attention design, Block Attention Residuals, and the Sol-Engine acceleration stack. The team says it trained unified 5B and 14B models from scratch on limited resources — 16 H100 nodes for the 5B model and 48 B200 nodes for the 14B model — and reports 84.30 VBench total, 3.2× faster DiT forward passes at 720p/60s, 13.06s for 720p/5s on a single H100, and 120× faster generation than Wan 2.2-A14B under the same setup.
Related event: NVIDIA Unveils SANA-Video 2.0 for Efficient Video Generation(2 posts)→
More from Multimodal
- Dev builds interactive 3D product experience with GPT-6 Astra + Hyper3D Rodin — nikola_mr64990 · 2026-09-11
- Using a finisher move on one mosquito with MiniMax H3 MAX — the bug survives — Hailuo_AI · 2026-09-11
- Skyfall GS Uses Flux to Refine Gaussian Splatting, Accepted at ECCV 2026 — ducha_aiki · 2026-09-11
- Lumara AI Film Festival Comes to NYC Oct 26, Top AI Filmmakers to Compete — 0xAllen_ · 2026-09-11
- Pterodactyl Detective: An AI-Generated Proof-of-Concept Trailer — PterodactylDetective · 2026-09-11
- Imperium Game Trailer Showcases AI Video Generation — keaslenyt · 2026-09-11