Swap any video subject with one image: Viggle-Animate (MiniMax H3 finetune) needs no pose detection, 3 inference steps
cocktailpeanut · x · 2026-09-15
A developer demos Viggle-Animate, a MiniMax H3 finetune for video subject swapping: extract the first frame, use GPT Image 2.5 to swap characters, then feed that single reference image plus the original video into the model — it syncs the reference to the video till the end. No pose detection, no SAM, no segmentation pipeline; multiple subjects can be swapped at once, and it runs in just 3 inference steps. The author calls it criminally underrated.
Related event: Maestro Local AI Studio and Video Character Swap Workflows Gain Attention(3 posts)→
More from Multimodal
- Leonardo breaks down Seedance 2.5 control techniques behind its AI short film — aziz4ai · 2026-09-15
- The SVG of a screenshot of a ChatGPT SVG meme shows how good GPT-6 Astra is — Angaisb_ · 2026-09-15
- Redditor finally gets a coherent video out of ComfyUI after weeks of struggling — Admirable_Yellow8170 · 2026-09-15
- Lyria 3.5 music demo stuns with vocals, laughter and movie-quote sampling — fofrAI · 2026-09-15
- Runway researcher: real-time generative world models are a key research push — c_valenzuelab · 2026-09-15
- Krea2 Turbo 2-Step Distill LoRA Hits New Checkpoint, Renders Up to 2048×2048 — TimeTruth2490 · 2026-09-15