MiniMax-H3 video inpainting ported to diffusers modular blocks: 6-step subject swap
linoy_tsaban · x · 2026-08-22
Linoy Tsaban ported MiniMax-H3's masked video and audio inpainting into Diffusers modular custom blocks (modular-diffusers), adapted from community ComfyUI workflows and released on Hugging Face under Apache-2.0.
In testing (with a turbo LoRA), a single reference photo plus 6 inference steps swaps the animal in a clip for the reference subject while the forest, snow, camera push, and original soundtrack stay untouched. The post includes a full Python example—loading blocks, passing mask/sourcevideo/sourceaudio and an image reference—plus a variant that skips the text encoder for split deployments.
More from coding & agent
- An Agent is 3 Layers: Business Logic, Harness, and Infrastructure — blaizedsouza · 2026-08-22
- Every Agent gains computer control, experiments reveal high token costs — every · 2026-08-22
- Grok Build 1.0.8 ships: concurrent subagents start faster, no more frozen sessions — kevinnbass · 2026-08-22
- GitHub Hit: A Massive Collection of Claude Custom Skills and Resources — tom_doerr · 2026-08-22
- Hands-on with Ox Alpha: Impressive Performance in Pi Harness — omarsar0 · 2026-08-22
- Replit Free Mode hailed as powerful ChatGPT with full cloud access — amasad · 2026-08-22