MiniMax H3 throughput and costs on RunPod benchmarks
Yasangas · reddit · 2026-08-22
A Reddit user inquires about real-world performance of the MiniMax H3 video model on RunPod. Key points include throughput on various GPU setups (4090, A100, H100), hourly output of 16:9 5-15s clips, and the use of INT8 quantization or Turbo LoRAs for optimization to estimate cost-per-video.
More from Infra
- H100 Shortage Driven by Power, Cooling, and Talent, Not Chip Fabrication — ingliguori · 2026-08-22
- Parsewave's work suggests a shift towards high-quality synthetic data in AI training — trashnash007 · 2026-08-22
- GLM 5.2 runs at 35t/s via DwarfStar mixed RAM/VRAM inference — antirez · 2026-08-22
- Data center pause may push AI jobs overseas amidst power bottleneck fears — Dan_Jeffries1 · 2026-08-22
- GLM-5.3 hits 21.4x speedup on RTX PRO 6000 via Kimi-Linear Decode kernel — teortaxesTex · 2026-08-22
- Why I'm traveling to India to build a dual-RTX 3090 rig for local LLMs — Ubunta · 2026-08-22