300 Renders Later: MiniMax H3 REF2VA Still Overshoots the Desired End Scene
bajungadustin · reddit · 2026-09-22
A Reddit user documents extensive struggles controlling MiniMax H3 REF2VA video generation: the camera keeps overshooting the intended final frame and returning to it. They tried First-Last, control video, depth control, FL2VA and roughly 300 renders, plus a massive structured prompt with subject definitions and keyframe specs. Notably, having ChatGPT write prompts with reference to the official prompt guide worked notably better. The post highlights current weaknesses in camera timing and endpoint control for video generation models.
More from Multimodal
- John Coogan declares "video is solved", teasing a video generation breakthrough — johncoogan · 2026-09-22
- Creative demo: your daily voice note becomes an AI-painted canvas with a matching melody — iamrobotbear · 2026-09-22
- Runway experiments with responsive generative video interfaces — DanScalco · 2026-09-22
- Redditor shows Qwen's image model can generate famous IP characters — Z3ROCOOL22 · 2026-09-22
- Right-Click Anything: SAM 3.1, Muse Spark and Muse Image combined into one clever demo — nikhilaravi · 2026-09-22
- A Video Prompt That Turns Gravity Upside Down: "The Other Floor" Full Prompt — umesh_ai · 2026-09-22