MiniMax H3's Reference Images Win Over an Open-Weights-Only Video Generation Holdout
Peregrine2976 · reddit · 2026-09-03
An open-weights-only creator says MiniMax H3 finally got him into video generation: Wan 2.2 never hit the quality or consistency he wanted, and the tools that did were closed-weight paid products.
- MMH3 is multimodal with vision, so you can feed reference images/character sheets directly — the feedback loop for teaching new concepts is far shorter than training LoRAs
- He doubts character/clothing/setting LoRAs will go obsolete, but expects far fewer will be needed
- Full workflow (Pastebin) and the reference sheet used (Imgur) are shared for reproduction
His stated motivation is ideological: tech that can be hacked apart and isn't beholden to anyone — same reason he runs Linux and Firefox.
More from Multimodal
- User lets Tavus' builder create a Pal of its own, and chaos ensues in this demo — Kyrannio · 2026-09-03
- Photoshop beta's Enhance Edge dramatically improves AI hair masking — rufusd · 2026-09-03
- Grok video model now supports up to 14 references in video generation — aziz4ai · 2026-09-03
- Meta's meta-models HF Org Surfaces Muse Glimmer 30B With 610k Downloads and New Papers — MaziyarPanahi · 2026-09-03
- Throwback: an accidental Mona Lisa generated with StyleGAN3 back in 2021 — Merzmensch · 2026-09-03
- Meta ships Muse Spark 1.3; side-by-side 3D viking figurine test vs 1.2 — alexandr_wang · 2026-09-03