Minimax Video Generation Test: Anime Character Consistency Remains Challenging

NeatUsed · reddit · 2026-08-11

The author discusses the difficulties encountered when using Minimax's reference-to-video (ref2vid) feature to generate anime clips. Similar to the previous Wan model, Minimax struggles significantly with maintaining facial consistency.

The specific pain point is that when trying to generate characters with distinct details (like the Sharingan), the model often ignores these key features and replaces them with generic anime elements. Even when using the reference image function, the model fails to accurately extract and apply specific details, causing visual fragmentation. The author seeks community advice on how to better control Minimax for character consistency.

Related event: Minimax Video Gen Struggles with Character Consistency: Community Explores H3 Workarounds(2 posts)→

Original post →

More from Multimodal

Multimodal channel →