ZICQ
中 Log in / Sign up
Newsroom LLMs & Foundation Models #Kandinsky #Audio-Video Generation #Diffusion Models #Multimodal AI #Open-Source Model

Kandinsky 6.0 Video Released: Revolutionizing Synchronized Audio-Video Generation

Avatar of Mr.Xu

By Mr.Xu

Published:

中文阅读 (Chinese) English Version

Summary:The Kandinsky team has released Kandinsky 6.0 Video, a family of foundation diffusion models for synchronized text-to-audio-video generation, including Kandinsky 6.0 Video Lite (3B parameters) and Kandinsky 6.0 Video Pro (29B parameters). These models generate 5-second video clips with synchronized 44 kHz audio, including lip-sync, in both text-to-audio-video (T2AV) and image-to-audio-video (I2AV) modes. A built-in super-resolution model boosts the output resolution to Full-HD (1920x1080). Lever


Source: Hugging Face Trending Papers (2026-10-04)

— END —

Tags: #Kandinsky #Audio-Video Generation #Diffusion Models #Multimodal AI #Open-Source Model

Community Comments

Loading live comments and annotations…