Kandinsky 6.0 Video: Foundation Models for Synchronized Video and Audio Generation Paper • 2610.05608 • Published 5 days ago • 155
NAMVIS: Next-Scale Autoregressive Multi-View Image Synthesis Paper • 2610.04722 • Published 6 days ago • 19
Qantara: Bridge-Flow Training for Multi-Paradigm JEPA Control Paper • 2607.04978 • Published Jul 6 • 8
Multi-LCB: Extending LiveCodeBench to Multiple Programming Languages Paper • 2606.20517 • Published Jun 18 • 61
RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments Paper • 2604.26067 • Published Apr 28 • 77
Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation Paper • 2511.14993 • Published Nov 19, 2025 • 236
G-CUT3R: Guided 3D Reconstruction with Camera and Depth Prior Integration Paper • 2508.11379 • Published Aug 15, 2025 • 12