Kandinsky 6.0 Video: Foundation Models for Synchronized Video and Audio Generation Paper β’ 2610.05608 β’ Published 3 days ago β’ 113
KVAE: Family of Tokenizers for Multimodal Generative Models Paper β’ 2608.05798 β’ Published Aug 6 β’ 31
Kandinsky 6.0 Video: Foundation Models for Synchronized Video and Audio Generation Paper β’ 2610.05608 β’ Published 3 days ago β’ 113
Your Transformer Can Hold Two Thoughts at Once: Evidence of Linear Superposition in LLMs Paper β’ 2609.29845 β’ Published 13 days ago β’ 104
KVAE-Audio Collection KVAE-Audio is a continuous full-band audio waveform autoencoder β’ 2 items β’ Updated 1 day ago β’ 7
Kandinsky 5.0 Video Pro Diffusers Collection Kandinsky 5.0 Video Pro is a 19B model that generates high-quality HD videos from English and Russian prompts with controllable camera motion. β’ 4 items β’ Updated 1 day ago β’ 15
AudioSAE: Towards Understanding of Audio-Processing Models with Sparse AutoEncoders Paper β’ 2602.05027 β’ Published Feb 4 β’ 63
T-pro 2.0: An Efficient Russian Hybrid-Reasoning Model and Playground Paper β’ 2512.10430 β’ Published Dec 11, 2025 β’ 121
Multimodal Evaluation of Russian-language Architectures Paper β’ 2511.15552 β’ Published Nov 19, 2025 β’ 79
Unveiling Intrinsic Dimension of Texts: from Academic Abstract to Creative Story Paper β’ 2511.15210 β’ Published Nov 19, 2025 β’ 91
Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation Paper β’ 2511.14993 β’ Published Nov 19, 2025 β’ 236