Benchmarking Minimax H3 Video Acceleration: 10s Video in 60s on a Single RTX 5090
nik_amaze · reddit · 2026-08-10
A developer on Reddit shared an aggressive local acceleration workflow for the Minimax H3 video generation model. Renting an RTX 5090 (32GB VRAM) with 56GB RAM from Vast.ai, the setup utilizes CUDA 13, SageAttn 2.2.0, and Triton 3.6.0 (installing Sol-Attn via git).
With all acceleration toggles enabled (excluding easyCache due to quality degradation), the int8 unpruned model with two image references generates a 10-second video at 1mp in 180s, and at 0.5mp in 60s. The creator shared the ComfyUI workflow and sought further optimization tips from the community.
More from Infra
- AI compute becomes strategic as tech giants pledge to build their own power infrastructure — bittingthembits · 2026-08-10
- DwarfStar Accelerates DeepSeek Inference with DFlash Speculative Decoding — antirez · 2026-08-10
- Neural AI Breakthrough: Memory Chip Reconstructs Human Cortex in Real Time — Dr_Alex_Crimi · 2026-08-10
- Cursor and Together AI Partner for Low-Latency AI Coding Inference — togethercompute · 2026-08-10
- SK hynix Pledges Additional Shareholder Return Measures for Q3 — toptickcrypto · 2026-08-10
- MiniMax H3 Test: Generates 30-Second Uncut Video in One Go — singularitynotnow · 2026-08-10