MiniMax Image-to-Video Test: 15s Clip Takes ~5 Minutes
rileygstaliger · reddit · 2026-08-06
The author shares a hands-on test of the MiniMax image-to-video model. Using a single starter image of multiple women and a prompt asking them to turn to the camera, the model achieved the desired result in just two takes.
- Generation Time: It took approximately 294 seconds (about 5 minutes) to generate a 15-second video.
- Workflow: Used a free ComfyUI workflow provided by Fox Fur Essence.
- Comparison: The author noted that trying to achieve the same result with LTX took forever, and MiniMax's quality is far superior.
Related event: MiniMax Image-to-Video Model Tested(2 posts)→
More from Multimodal
- Minimax H3 video generation stuck in 'uncanny valley', dev says — mattshumer_ · 2026-08-06
- Grok Imagine to Introduce Keyframe Support and Upgraded Image Model — chaitu · 2026-08-06
- Nano Banana Prompt Workflow: Turning Any Logo into Realistic Impasto Art — azed_ai · 2026-08-06
- Running MiniMax H3 Locally on a 6-Year-Old GPU to Generate 76 Diverse Animation Clips — PetersOdyssey · 2026-08-06
- Gemini Image Gen Combined with GPT Coding Easily Creates AI Visual Heroes — RichardsonDx · 2026-08-06
- Testing MiniMax H3: Generating Various English Accents — wikid24 · 2026-08-06