MiniMax H3 Ref2V pain points: 15-min renders, broken audio, wrong durations
yeah280 · reddit · 2026-09-07
A user testing MiniMax H3's Ref2V in ComfyUI for character/background replacement reports three problems: original audio isn't preserved, a 3-second video takes 15 minutes to generate, and output duration doesn't match the source. Their setup uses refvideo0/refaudio0 nodes, 20 steps at 24FPS, and a prompt asking to keep motion, camera, timing and audio while removing subtitles. Workflow JSON attached, seeking help on the wiring.
More from Multimodal
- Turning AI-generated pixel sprites into animations with Sprite Fusion — HugoDuprez · 2026-09-07
- Grok, ChatGPT and Gemini All Failed at Making a Collage of His Book Covers — pickover · 2026-09-07
- Designer concedes GPT-6 Astra 'nearly unbeatable' after it designed full Figma pages via Computer Use — deedydas · 2026-09-07
- GPT-6 Astra Turns Van Gogh's Starry Night Into a Walkable 3D World — emmanuelvivier · 2026-09-07
- GPT-6 Astra's Iterative 3D Modeling Demos Impress Early Users — Cklly2004 · 2026-09-07
- AI Short Film 'Mechanical Love' Made with Seedance 2.5 and GPT Image 2.0 — HashemGhaili · 2026-09-07