DEFINE: Exemplar-Guided Accent Control for Zero-Shot TTS Paper • 2609.32777 • Published 12 days ago • 35
Kandinsky 6.0 Video: Foundation Models for Synchronized Video and Audio Generation Paper • 2610.05608 • Published 8 days ago • 166
Running on Zero Agents 9 Kandinsky 6.0 Pro Distill 5s 🔥 9 10-step Kandinsky 6.0 Pro text/image to video with audio
kandinskylab/Kandinsky-6.0-Lite-pretrain-5s-Diffusers Image-to-Video • 3B • Updated 4 days ago • 101 • 11
kandinskylab/Kandinsky-6.0-Pro-pretrain-5s-Diffusers Image-to-Video • 30B • Updated 4 days ago • 94 • 10
kandinskylab/Kandinsky-6.0-Lite-distill-5s-Diffusers Image-to-Video • 3B • Updated 4 days ago • 370 • 22
kandinskylab/Kandinsky-6.0-Lite-5s-Diffusers Image-to-Video • 3B • Updated 4 days ago • 255 • 29
kandinskylab/Kandinsky-6.0-Pro-distill-5s-Diffusers Image-to-Video • 30B • Updated 4 days ago • 406 • 28
kandinskylab/Kandinsky-6.0-Pro-5s-Diffusers Image-to-Video • 30B • Updated 4 days ago • 277 • 51
Kandinsky 6.0 Video: Foundation Models for Synchronized Video and Audio Generation Paper • 2610.05608 • Published 8 days ago • 166
KVAE-Audio Collection KVAE-Audio is a continuous full-band audio waveform autoencoder • 2 items • Updated 6 days ago • 7
KVAE: Family of Tokenizers for Multimodal Generative Models Paper • 2608.05798 • Published Aug 6 • 32
KVAE: Family of Tokenizers for Multimodal Generative Models Paper • 2608.05798 • Published Aug 6 • 32
KVAE-Audio Collection KVAE-Audio is a continuous full-band audio waveform autoencoder • 2 items • Updated 6 days ago • 7