VideoDeltaNet open-sources hybrid attention that speeds up MiniMax H3 video generation up to 90x
realmrfakename · x · 2026-09-03
OpenVDN released VDN-H3, a hybrid-attention video model built on MiniMax H3 claiming up to 90x speedup with near-lossless quality, outperforming FastH3.
- Hybrid architecture: a frame-wise linear attention branch provides efficiency while a softmax branch preserves backbone visual quality and consistency.
- Plug-and-play: implemented as a new linear attention branch plus two small LoRA adapters merged at inference, leaving backbone weights untouched.
- Speed: generates a 14.4-second clip in 11.23 seconds on 8x B200 with 8 denoising steps — faster than real-time playback.
- Fully open source: 82 GB of weights (including the 50-step distillation source and 8-step turbo model), an optimized inference stack, and training code are all released.
Related event: VideoDeltaNet Speeds Open Video Generation Up to 90x(4 posts)→
More from Infra
- Inference Engineering Is Just a Recipe: vLLM/SGLang, Replicas, Cache-Aware Routing — GabGarrett · 2026-09-03
- Databricks pitches agent-native data infrastructure, Lakebase Postgres at VLDB 2026 — matei_zaharia · 2026-09-03
- Lablup, Maker of GPU Orchestrator Backend.AI, Joins PyTorch Foundation as Silver Member — PyTorch · 2026-09-03
- Cursor cloud agents can now run on your own infrastructure, Mac Minis included — mattyp · 2026-09-03
- Analyst: NVIDIA Could Become Intel Foundry's 'Customer Zero' as a Second Source Beyond TSMC — BenBajarin · 2026-09-03
- Investors bullish on Meta as Muse Spark 1.3 pricing undercuts frontier rivals — Scobleizer · 2026-09-03