Running MiniMax H3 on RTX 5090: Video-to-Video Generation Takes 20 Minutes
Chaztle · reddit · 2026-08-09
A user reported a significant processing time discrepancy when running the MiniMax H3 model locally on an RTX 5090. While text-to-video and image-to-video speeds are tolerable (taking about 2m 30s for a 720p, 5s, 8-step video), video-to-video tasks take a massive 20 minutes. The poster is looking for insights into why this bottleneck occurs and how to speed it up.
Related event: RTX 5090 Struggles with MiniMax H3 Local Inference(2 posts)→
More from Infra
- DeepSeek Local Deployment: Troubleshooting Severe Speed Drop with Speculative Decoding — Easy_Werewolf7903 · 2026-08-09
- Fixing Black Video Outputs with MiniMax H3 on AMD GPUs — Present-Guitar-3967 · 2026-08-09
- Enabling PCIe P2P on Consumer Nvidia GPUs Boosts LLM Throughput by 25% — BidonPomoev · 2026-08-09
- Meituan's LongCat 2.0: Fully Trained and Inferenced on Chinese ASICs — bycloud · 2026-08-09
- Which 4-bit Quant is Best for MLX? Comparing Mainstream Options — True_Tangerine_4706 · 2026-08-09
- Amazon's Planned Texas Data Center Power Plant Could Become Top US Climate Polluter — TechCrunch AI · 2026-08-09