Testing MiniMax H3 Full-Reference Workflow: Consistent Characters and Lip-Sync
Artefact_Design · reddit · 2026-08-09
A creator shared their workflow for producing a music video using the MiniMax H3 model.
- Inputs: Character reference sheets, location references, and an isolated vocal stem for lip-syncing (no music in the audio input).
- Results: Everything else was generated by the model. The creator noted satisfaction with the consistency and lip-sync holding up across different scenes and characters.
- Prompting: They are still dialing in the prompt structure shot by shot to optimize the output.
Related event: Creators Test MiniMax H3 for AI Music Video Generation(3 posts)→
More from Multimodal
- Gemini Multimodal Demo: Building a Tiny Luxury Farmhouse via Prompt Workflow — bennash · 2026-08-09
- ComfyUI Video Edition to Support Minimax for Video Chaining and Motion Reference — shootthesound · 2026-08-09
- Temporary Custom Node Fix for ComfyUI Widget Width Bug Released — Budget_Competition77 · 2026-08-09
- AI-Generated Retro Japanese Anime Opening (Naruto Fan Edit) — GamerVick · 2026-08-09
- Recreating Retro Anime OP with MiniMax H3: 15-min Generation on RTX 5090 — GamerVick · 2026-08-09
- MiniMax H3 ComfyUI LoRAs Spotted in Kijai Repository — FlatwormMean1690 · 2026-08-09