Building characters straight from MiniMax H3 T2V + R2VA, no more SDXL anchors
SIR_NVAX_A_LOT · reddit · 2026-08-17
The author ditched the old SDXL+KREA2 anchor-image pipeline and went straight to MiniMax H3's T2V plus R2VA for character generation: rooftop actors were incepted via T2V, cleaned up with V2V, then detailed with an anime reference for the pilot's plugsuit. The cockpit scene was generated purely from a prompt, with the HUD/mecha-kaiju referenced from a 15s battle render. They also experimented with transferring a Kaiju's look onto a woman's armor. One caveat: R2VA sometimes makes actors wider/squished, so FL2VA is safer if you don't want re-envisioning.
More from Multimodal
- Prediction market: Alibaba has 13% chance to top AI models by year-end — Polymarket · 2026-08-17
- Alibaba launches AI music generation model "HappyShrimp" — Polymarket · 2026-08-17
- Seedance 2.0 4K vs. Seedance 2.5 1080p video quality comparison — DavidmComfort · 2026-08-17
- Short film 'Reincarnation' made with Flux 3 — Tadeo111 · 2026-08-17
- An entire music video made with MiniMax H3 visuals, Suno music, and DaVinci Resolve — la_art · 2026-08-17
- LTX-2.5 vs MiniMax H3 i2v on an RTX 5090: 1080p vs 1344×768 is what 32GB fits — chanteuse_blondinett · 2026-08-17