Testing Minimax H3 Multi-Shot Prompting for Image-to-Video
call-lee-free · reddit · 2026-08-19
The author shares a detailed test of Minimax H3's image-to-video capabilities using multi-shot prompting. By designing a prompt sequence with timestamps and dialogue, a 30-second continuous video was generated. The post details the full prompt structure, including scene descriptions, character lines (with specified British accents), and camera cuts. Limited by hardware (RTX 4070 Super), the render was capped at 0.4MP resolution and took 173 minutes to complete.
More from Multimodal
- From character creation to stories: AI influencers streamline anime workflows — aftahi_ai · 2026-08-20
- Grok demo: Accurate language accent and long 1080p video generation — elonmusk · 2026-08-20
- int21.ai demos swarm-built inference engines for video, music, and speech — bingxu_ · 2026-08-20
- AI Tool GeoSpy Locates Photos from Pixels with Meter-Level Accuracy — saibharadwaj · 2026-08-20
- Runway Gen-2 Update: 1080p Support, 50 References, 30s Generation via API — tlakomy · 2026-08-20
- Digital Sculpting: Creating the Thesis Rock with Rendering Magic — every · 2026-08-20