Open-Source LynnReal-Omni: One 32B Model for Video Gen, Editing and Restoration, 377ms on H100
AgeNo5351 · reddit · 2026-09-16
LynnReal-Omni is out: a 32B shared multimodal diffusion transformer built on the MiniMax H3 architecture, with open weights and ComfyUI nodes.
One model, many video tasks
- Text-to-video, image-to-video, human/hand pose-guided generation, structural control, omni-reference, style transfer, and video editing;
- Restores degraded video (reportedly also repairs accumulated-error artifacts);
- Streaming long-video generation;
- Accepts heterogeneous inputs (appearance refs, editable 3D renders, game recordings) so an agent can compose visual conditions in one model;
- Four-step fast generation.
Flash for real-time rendering
- A 27B three-step variant with lightweight VAE decoder;
- On a single H100, 22-frame 540p generation + decoding takes 843ms (Standard) or 377ms (Flash), laying groundwork for real-time streaming video.
Related event: LynnReal-Omni: Open-Source Unified Video Generation Framework(2 posts)→
More from Multimodal
- Grok turns out great at 3D: Blender MCP workflow yields video in ~2 hours — techartist_ · 2026-09-16
- Premiere adds "Sync to video" button for auto-aligned AI sound effects — justin_salamon · 2026-09-16
- After a year in Firefly, Adobe sound generation lands in Premiere — justin_salamon · 2026-09-16
- Solo creator makes 4-min AI short film in 8 days with Higgsfield — philrox_ · 2026-09-16
- AI-made promo film for Shanghai Light Festival features an orange cat in naked-eye 3D — ring_hyacinth · 2026-09-16
- AI-generated 'Dune: Bene Gesserit Origin' cinematic short film, director's cut — Technical_Bite_4488 · 2026-09-16