Swapping Complex Fight Scene Backgrounds with MiniMax Using Only 2 Photos
illumi_nation25 · reddit · 2026-08-14
The author shares an advanced workflow for high-difficulty video background replacement using the MiniMax video model (MMH3 REF2V). The experiment shows that by using just 2 real-life photos as references, the model can successfully replace the background of a 12-second complex fight scene while preserving the original motion, lighting, and camera trajectory.
Key Workflow & Insights:
- Hardware/Software: Run on RTX 5090, utilizing DaSiWa MiniMax H3 Workflows, with final editing in DaVinci Resolve.
- Chunking Strategy: Since the model struggles to match long choreography perfectly, the author split the 12s video into two 6s segments, generated them separately, and stitched the best parts.
- Prompt Structure: The post shares a detailed structured prompt template. It explicitly defines retention and replacement rules for video elements (camera path, action, audio) and environment elements (target background), achieving precise visual control.
More from Multimodal
- PROJECTIFY: Projecting ComfyUI-Generated Images Onto 3D Models — AccordingInspector58 · 2026-08-14
- SandAI Unveils MAGI-2 Preview: 114B Parameter Open-Source Video Model — zephyr_z9 · 2026-08-14
- Creating Ugly-Cute Blind Box Characters with AI Proves Hilariously Fun — CommitteeMedical5449 · 2026-08-14
- OnSolo Integrates Seedance 2.5 for Single-Prompt Video Generation — SucceededMind · 2026-08-14
- AI Video Generation Test: Pure Prompts Struggle with Precise Dance Choreography — R34vspec · 2026-08-14
- Spotting AI Fakes: Expert Details Bizarre Character Stroke Artifacts — brianryhuang · 2026-08-14