MiniMax-H3 Multimodal Model Surfaces, Runs on ComfyUI with a Single 24GB GPU
reeight · reddit · 2026-08-03
The MiniMax-H3 model (RunningHub version) is now available on Hugging Face. It supports generating video+audio from image/audio/video references (e.g., T2VA, Ref2VA).
According to community feedback, a corresponding ComfyUI custom node repository is already available, with a local deployment threshold of a single 24GB consumer-grade GPU.
Related event: ComfyUI Day-0 Support for MiniMax H3 Cuts VRAM by 66%, Runs on RTX 3060(8 posts)→
More from Multimodal
- Developer Creates 'Song of the Sirens' Concept Art Using Grok — bennash · 2026-08-03
- xAI Updates Grok Imagine 1.5, Wowing Users with Image Generation Upgrades — minchoi · 2026-08-03
- Seedance Video Model Excels at Crisp Vector-Style Graphics — bennash · 2026-08-03
- Four Years of AI Image Generation Evolution: Same Prompt, Worlds Apart — DreamFly_13 · 2026-08-03
- Motion Skill Turns Claude into a Video Team: Generate Launch Videos from a URL — Scobleizer · 2026-08-03
- MiniMax Launches H3: Native 2K Video with Stereo Audio Generation — Mobile-Pumpkin7944 · 2026-08-03