Testing MiniMax H3 R2V Character Swap Falls Short on Action Matching

CtrlAltDefeat- · reddit · 2026-08-08

A user reported significant struggles when using the Reference to Video (R2V) feature of the MiniMax H3 model. When attempting to swap a character in an input video using reference images, the model often retains the original character or struggles to match complex actions accurately. However, generating purely from reference images without an input video works well, highlighting current limitations in video-to-video action control.

Related event: MiniMax H3 R2V Video Generation Struggles with Face Consistency(2 posts)→

Original post →

More from Multimodal

Multimodal channel →