MiniMax H3 Tested: Natural Video and Voice from Blurry Photos

VoidAsuka · x · 2026-07-31

A user tested the MiniMax H3 model, successfully generating a video using two blurry photos from two years ago. The test revealed natural motion and high-fidelity voice generation that closely matched the user's real speech.

The author noted the rapid progression in video generation, which can now handle high-fidelity audio, accurate text rendering, and editable video simultaneously—a massive leap from the early days of GPT-4o and Sora.

Related event: MiniMax H3 Video Model Impresses in Comprehensive Tests(57 posts)→

Original post →

More from Multimodal

Multimodal channel →