MiniMax H3 ecosystem roundup: multi-GPU recipes, 24GB VRAM runtime, turbo LoRAs
optimisticalish · reddit · 2026-09-12
A Reddit roundup of the MiniMax H3 ecosystem:
- Inference: RunningHub's H3 Lightning multi-GPU acceleration recipe; FastVideo's FastH3-Preview-v0.2 and FastH3-Live 1.2.0.
- Low-end: New VideoDeltaNet weights plus a ComfyUI runtime that runs on 8×B200 or a single 24GB card.
- Training & speed: Experimental DMD turbo 8-step LoRAs; an Image-Training-Adapter aiming at LoRA concept training without degrading video knowledge.
- Workflows: ComfyUI-MM-1Frame single-frame extractor (auto-picks best of five frames, accepts an OpenPose reference with a one-word pose trick to avoid stick-figure artifacts); a no-RefMod multi-image reference workflow; a Portuguese 3D camera-control widget; and Depthcat for depth blockouts (Apache-2.0, 111MB extra model).
More from Infra
- Instinct may burn $100M+ a year in tokens, and open-weight models aren't actually cheaper — ivan_bezdomny · 2026-09-12
- Curie: a from-scratch 17B model designed to run from SSD, 33 tokens/s on one CPU core — Just_Vugg_PolyMCP · 2026-09-12
- LMStudio now accepts llama.cpp overrides; --yarn-attn-factor 1.2 may boost creativity — Extraaltodeus · 2026-09-12
- Analyst details Apple's S11, A20 and M6 silicon: new packaging, cooling and on-device AI designs — BenBajarin · 2026-09-12
- Custom Silicon 3.0: market shifts to program responsibility as agentic AI enters cyber defense — BenBajarin · 2026-09-12
- Single 3090 + 32GB RAM: squeezing max fidelity out of local Qwen3.8-27B — Certain_Yam_5824 · 2026-09-12