User reports MiniMAX H3 struggles with facial consistency in image-to-video

max4634 · reddit · 2026-09-02

A user testing MiniMAX H3 found that its image-to-video feature fails to maintain facial consistency. When generating a talking scene from a photo, the face drifts and morphs into a different person once movement starts, despite keeping the general vibe. Comparison tests suggest it struggles more with facial stability than the LTX model.

Original post →

More from Multimodal

Multimodal channel →