DEFINE: Exemplar-Guided Accent Control for Zero-Shot TTS Paper • 2609.32777 • Published 8 days ago • 33
Kandinsky 6.0 Video: Foundation Models for Synchronized Video and Audio Generation Paper • 2610.05608 • Published 4 days ago • 137
Running on Zero Agents 7 Kandinsky 6.0 Pro Distill 5s 🔥 7 10-step Kandinsky 6.0 Pro text/image to video with audio
kandinskylab/Kandinsky-6.0-Lite-pretrain-5s-Diffusers Image-to-Video • 3B • Updated about 21 hours ago • 80 • 8
kandinskylab/Kandinsky-6.0-Pro-pretrain-5s-Diffusers Image-to-Video • 30B • Updated about 21 hours ago • 82 • 8
kandinskylab/Kandinsky-6.0-Lite-distill-5s-Diffusers Image-to-Video • 3B • Updated about 21 hours ago • 171 • 16
kandinskylab/Kandinsky-6.0-Lite-5s-Diffusers Image-to-Video • 3B • Updated about 21 hours ago • 128 • 24
kandinskylab/Kandinsky-6.0-Pro-distill-5s-Diffusers Image-to-Video • 30B • Updated about 21 hours ago • 279 • 22
kandinskylab/Kandinsky-6.0-Pro-5s-Diffusers Image-to-Video • 30B • Updated about 21 hours ago • 216 • 41
Kandinsky 6.0 Video: Foundation Models for Synchronized Video and Audio Generation Paper • 2610.05608 • Published 4 days ago • 137
KVAE-Audio Collection KVAE-Audio is a continuous full-band audio waveform autoencoder • 2 items • Updated 2 days ago • 7
KVAE: Family of Tokenizers for Multimodal Generative Models Paper • 2608.05798 • Published Aug 6 • 31
KVAE: Family of Tokenizers for Multimodal Generative Models Paper • 2608.05798 • Published Aug 6 • 31
KVAE-Audio Collection KVAE-Audio is a continuous full-band audio waveform autoencoder • 2 items • Updated 2 days ago • 7