Running MiniMax H3 Locally: Extremely High VRAM and RAM Usage Reported
Full_Astronomer_5438 · reddit · 2026-08-03
Reddit users report extremely high resource consumption when running the MiniMax H3 model locally in ComfyUI. The model causes severe inference spikes and drives GPU temperatures to unusually high levels.
Even with a pruned INT8 quantization and NVFP4 text encoder, the model quickly consumes 10-30GB of system page file (virtual memory) depending on the video length (5-15 seconds).
More from Infra
- Run DeepSeek V4-Flash with 1M Context Locally on Dual RTX PRO 6000 — dee_hw · 2026-08-03
- Local AI on a $1500 Budget: Handling Private Data Efficiently — KookyThought · 2026-08-03
- AMD CEO Predicts $2T Computing Market by 2030, Driven by $1.4T AI Accelerators — Beth_Kindig · 2026-08-03
- Hacking Liquid Cooling for AMD MI250X: A $1400 AI Accelerator Steal? — MLDataScientist · 2026-08-03
- Cloudflare Restructures Cloud Agent Architecture: Container Sandboxes On-Demand — irvinebroque · 2026-08-03
- Cloudflare Launches Billable Usage API for Programmatic Cost Visibility — ritakozlov · 2026-08-03