MiniMax H3 Dropped: T2V, I2V, and Reference-to-Video with Audio
multimodalart · x · 2026-08-03
MiniMax has officially released the H3 video generation model. With 33 billion parameters, it supports text-to-video, image-to-video, and reference-to-video capabilities, complete with synchronized audio generation.
The model weights are now available on Hugging Face. It has been integrated into the 🧨 Diffusers and ComfyUI ecosystems, making it ready for deployment and experimentation on consumer-grade GPUs.
More from Multimodal
- xAI Updates Grok Imagine 1.5, Wowing Users with Image Generation Upgrades — minchoi · 2026-08-03
- Seedance Video Model Excels at Crisp Vector-Style Graphics — bennash · 2026-08-03
- MiniMax H3 Gets Day 0 Support in SGLang, Runs Locally on Dual RTX 5090s — ying11231 · 2026-08-03
- Motion Skill Turns Claude into a Video Team: Generate Launch Videos from a URL — Scobleizer · 2026-08-03
- MiniMax Launches H3: Native 2K Video with Stereo Audio Generation — Mobile-Pumpkin7944 · 2026-08-03
- ComfyUI Adds Day 0 Support for MiniMax Video Model, Slashing VRAM by 66% for RTX 3060 — crystal_alpine · 2026-08-03