Fixing MiniMax H3 Local OOMs: 3090 Runs 30s Video with Tweaked Args
knoll_gallagher · reddit · 2026-08-05
While running the MiniMax H3 model locally on an RTX 3090 (64GB), a developer discovered a surprising cause for OOM (Out of Memory) errors: GPU demands being too low, which triggered dynamic memory allocation issues.
By tweaking ComfyUI launch arguments (like --vram-headroom 3), the author successfully generated a 30-second video at 0.2 megapixels. Using int8 and kj sage optimizations, a single video takes about 16 minutes. The author recommends keeping the H3 environment separate from the main ComfyUI instance due to its unique scheduling behavior.
Related event: Running MiniMax H3 on a Single RTX 3090(2 posts)→
More from Infra
- NSF Launches $100M Program for Regional AI Infrastructure Hubs — mkratsios47 · 2026-08-05
- Chutes AI Enforces TEE Verification: 8x RTX 5090s Beat Pro GPUs at 65% Lower Cost — markjeffrey · 2026-08-05
- engyai Launches Cheapest Kimi K3 API on OpenRouter, Cutting Costs by 50% — const_reborn · 2026-08-05
- LiquidAI's LFM2.5-2.6B Hits 82 tok/s Decode on Mac with 128K Context — helloiamleonie · 2026-08-05
- ai& Partners with Voltaiq for Battery Storage in Japanese AI Data Centers — DavidBennett__ · 2026-08-05
- Best Local LLMs for Coding on a 128GB Mac? — Electronic_Back1502 · 2026-08-05