MiniMax H3 offers multimodal references and granular control
thetripathi58 · x · 2026-08-16
MiniMax H3 demonstrates granular control capabilities:
- Multimodal Inputs: Accepts up to 9 images, 3 videos, and 3 audio references in a single prompt.
- Consistent Generation: Generates up to 15 seconds of footage with stable camera movements and character actions.
- Instruction-based Editing: Allows for editing via text instructions without regenerating the entire clip, providing a true directing experience.
More from Multimodal
- RTX 5090 turns full manga chapter into animated book in ~3 hours, fully automated and local — Ill-Ant-9489 · 2026-08-16
- Ghost Genesis: Machine Learning Meets Medieval Scripture — misovalko · 2026-08-16
- ReSplat: Learning Recurrent Gaussian Splatting accepted to ECCV 2026 Oral — rsasaki0109 · 2026-08-16
- MiniMax H3 image-trained LoRA tested for video generation — bdsqlsz · 2026-08-16
- Midjourney Prompt Showcases Convex Mirror Luxury Contrast, AI Art Creativity — tisch_eins · 2026-08-16
- Creator builds full Horizon Zero Dawn trailer using AI, no studio needed — heypearlai · 2026-08-16