MiniMax H3 ecosystem roundup: quantization, multi-GPU inference, LoRA tooling and workflows in one week
Just_Lingonberry_352 · reddit · 2026-09-19
A Reddit roundup aggregating the rapid growth of the MiniMax H3 video-generation ecosystem from Sept 10–18, 2026:
- Quantization & speed: W4A8 quantized Fun ControlNet-Union (2.13GB→1.45GB); RunningHub's multi-GPU inference recipe; VideoDeltaNet demoed on 8×B200 with a 24GB-VRAM consumer runtime; experimental 8-step DMD Turbo LoRAs.
- Training: Fizgig LoRA trainer 5.5.0 enables weight averaging by default and can inspect what each of H3's 52 LoRA blocks affects (motion, faces, audio).
- Workflows: a local pixel-perfect pixel-art animation guide; Spherical VAE decoding for 360°/VR video; a visual RefMod browser for ComfyUI; ComfyUI 0.35.0 with MiniMax updates; MiniMax Music Production Toolkit 2.1.1.
More from Infra
- Jev scales on Modal, one of only 4 subprocessors listed, to meet surging demand — AAAzzam · 2026-09-19
- Tuning Qwen3 27B as a coding agent on 2x3090s cuts turn latency from 28s to 7s — bolts98 · 2026-09-19
- AgentZip: memory compression for parallel agent sandboxes cuts memory up to 8.7x — rohanpaul_ai · 2026-09-19
- US products quietly build on Chinese open-weight models as one firm cuts spend by ~100x — generativist · 2026-09-19
- Turbines sold out through 2030: 25 categories of datacenter power gear face multi-year backlogs — CatAstro_Piyush · 2026-09-19
- Why OpenAI and Anthropic are suddenly buying tiny 20-30MW data centers — abhiadesai · 2026-09-19