Recreating videos in MiniMax H3: a 4-step frame-extraction workflow
Negative-Whereas3307 · reddit · 2026-09-18
A Reddit user demonstrates recreating a Mario talking-head video in MiniMax H3, arguing that "make something similar" leaves too much unspecified. The 4-step workflow (from Hypit, also doable manually):
- Extract reference frames at points where expression, pose, or on-screen content changes.
- Break down the scene: describe framing, setting, delivery, and overlays separately (presenter, ranking numbers, airplane card are independent parts).
- Decide what to keep: specify what H3 should recreate vs. what can change, keeping performance instructions separate from graphics instructions.
- Compare both versions side by side and fix specific gaps in the next pass.
The demo keeps the talking-head setup and numbered list but misses the airplane card — exactly the kind of gap the breakdown step is meant to catch.
More from Multimodal
- LightOnOCR-2-1B: 1B OCR model beats rivals 9x its size, 493K pages/day on one GPU — thisguyknowsai · 2026-09-18
- 2D animations in 15s: full GPT spritesheet + H3 Max Magnific workflow shared — techhalla · 2026-09-18
- KITScenes Multimodal Dataset Launches to Fill the Data Gap for 3D Foundation Models — abursuc · 2026-09-18
- Qwen-Image-2.1 specs leak: 7B DiT, Qwen3-8B encoder, 56GB VRAM for 4K — bdsqlsz · 2026-09-18
- AI artist 'Will' drops a new song daily, with persistent memory and personality — kun101 · 2026-09-18
- AI-made star IP car ads cut production from 2 months to 2 weeks using ByteDance Seedance and AMD compute — 机器之心 · 2026-09-18