Minimax Video Face-Swap Fails Past 5s? Prompting & Param Tips
Mediocre-Toe3212 · reddit · 2026-08-06
While running Minimax video generation in ComfyUI for reference-based face swaps, the author encountered a matching failure: the model loses facial lock when generation exceeds 5 seconds.
Through extensive testing, several workarounds were identified:
- Step Count: 12-15 steps work best. The default 20 steps cause the model to overthink and fail the match.
- Ref Image Size: Setting to MAX yields better results than match.
- Prompting: Describing specific actions (e.g., "attacking with knives") tends to alter the original video scene. Sticking to subject and static environment descriptions (e.g., "surrounded by men with knives") improves success rates.
The author still cannot reliably generate swaps beyond 5-6 seconds and is seeking further community advice.
More from Multimodal
- Testing MiniMax H3 native ComfyUI on RTX 3060 12GB: Setup guide — Creepy-Fault6977 · 2026-08-06
- Generating with MiniMax H3: If Darwin Presented Evolution Theory in 2026 — HeyZoyaKhan · 2026-08-06
- 12-min generation on a single GPU: MiniMax H3 local uncensored test — 3Dave_ · 2026-08-06
- MiniMax H3 Video Generation Bug: Garbled and Incomprehensible Speech — Alex_the_tiktock · 2026-08-06
- MiniMax H3 Model Tested: Excels at Anime Generation — cloneofsimo · 2026-08-06
- Driving ComfyUI with Claude Code: Ideogram 4.0 Workflow Tested — nark0se · 2026-08-06