MIT releases open-weight Kandinsky video+audio model in 6 variants with Diffusers day-0 support
RisingSayak · x · 2026-10-06
MIT has released Kandinsky, an open-weight video model with joint audio generation. It ships in 6 variants balancing speed, memory, and quality, plus two super-resolution checkpoints. Hugging Face Diffusers offers day-0 integration, so the model is usable on release day.
Related event: Kandinsky 6.0 Video Open-Sources Audio-Video Generation Under MIT License(4 posts)→
More from Multimodal
- Google rolls out Nano Banana 2.1 with 4K output at $0.076 per image — testingcatalog · 2026-10-06
- 85mm Editorial Portrait Prompt Template That Preserves Facial Identity — aziz4ai · 2026-10-06
- Solaya turns a 3-minute iPhone scan into a photorealistic 3D digital twin in under an hour — willeastcott · 2026-10-06
- Dev makes promo video with Opus and fframes, skipping hours of After Effects work — kevinkern · 2026-10-06
- Viral prompt recipe makes GPT image models shoot 'unpublished' photojournalism — techhalla · 2026-10-06
- Fun Seedance video demo shared with the exact prompts to recreate it — techhalla · 2026-10-06