Kandinsky 6.0 Video launches: open foundation models for synced video and audio generation

q5sys · reddit · 2026-10-08

Kandinsky Lab has released Kandinsky 6.0 Video, open foundation models for synchronized video and audio generation, with weights available on Hugging Face in diffusers format. They also shipped a companion video upscaling (VSR) model. The poster found v5 underwhelming and hopes v6 improves.

Related event: Kandinsky Lab Open-Sources Kandinsky 6.0 Video with Joint Audio-Video Generation under MIT License(7 posts)→

Original post →

More from Multimodal

Multimodal channel →