Kandinsky 6.0 Video Released: Revolutionizing Synchronized Audio-Video Generation
By Mr.Xu
Published:
Summary:The Kandinsky team has released Kandinsky 6.0 Video, a family of foundation diffusion models for synchronized text-to-audio-video generation, including Kandinsky 6.0 Video Lite (3B parameters) and Kandinsky 6.0 Video Pro (29B parameters). These models generate 5-second video clips with synchronized 44 kHz audio, including lip-sync, in both text-to-audio-video (T2AV) and image-to-audio-video (I2AV) modes. A built-in super-resolution model boosts the output resolution to Full-HD (1920x1080). Lever
— END —Source: Hugging Face Trending Papers (2026-10-04)
Tags: #Kandinsky #Audio-Video Generation #Diffusion Models #Multimodal AI #Open-Source Model
Community Comments