Timeline as a slot axis: LoRA turns MiniMax H3 into one-photo 5-view turnaround
linoy_tsaban · x · 2026-09-18
A comparison of two camera-view-control approaches for MiniMax H3:
- Turnaround LoRA: repurposes H3's video timeline as a "slot axis" — five single-image latents are packed into a 5-frame video latent with stretched RoPE positions and jointly denoised, so the transformer's temporal consistency becomes cross-view identity. One reference image plus one instruction yields five coherent rotated views in 10s (512²) to 57s (1024²) on one GPU. Though trained only on H3-generated subjects, identity transfer to real photos works out-of-distribution. Proof of "video brain, image body": applied to normal video generation, motion breaks — windmill blades render as superimposed discrete positions, since the LoRA learned time = poses.
- Meridian: finetunes H3 on geometry-warped references instead.
Weights are open on Hugging Face (s1500 for best rotation geometry, s400 for better instruction following).
More from Multimodal
- Self-proclaimed ChatGPT co-inventor launches Jev, claiming 20-200x speed and 40-400x cost gains — round · 2026-09-19
- Converting video to 120fps with a single prompt via Runway MCP — tlakomy · 2026-09-19
- Runway MCP turns video to 120fps with a single prompt — tlakomy · 2026-09-19
- From chemist to AI artist: inside one of the first AI teams on a TV series — Loo_Atreides · 2026-09-19
- Nautilo launches co-creation harness for AI: generate and edit video on one timeline — Dan_Jeffries1 · 2026-09-19
- Tencent's Unreleased Hunyuan 3.5 Image Model Spotted in Early Access, Pitted Against GPT Image-2 — FellMentKE · 2026-09-19