SoL-Refiner: Speed-of-Light One-Step Refinement for High-Resolution Video Paper • 2609.37969 • Published 9 days ago • 41
EVO-WAM: Evolving World Action Models through Video-Action Verification Paper • 2609.38057 • Published 9 days ago • 48
WorldLine: Action-Driven Visual Simulation for Robotic Manipulation Paper • 2609.38059 • Published 9 days ago • 32
Unfold The World: Factorize 4D Properties in Reinforcing Spatial Reasoning Paper • 2609.03729 • Published Sep 3 • 10
ZimaBlue: Evolving Generalizable World Action Models through Scalable Video Pre-training Paper • 2609.00188 • Published Aug 31 • 54
JoyAI-Video-Edit: Real-Time Open-Ended Video Editing with Autoregressive Diffusion Paper • 2608.03974 • Published Aug 4 • 107
TextLDM: Language Modeling with Continuous Latent Diffusion Paper • 2605.07748 • Published May 8 • 25
OpenSpatial: A Principled Data Engine for Empowering Spatial Intelligence Paper • 2604.07296 • Published Apr 8 • 39
Vision-Language-Vision Auto-Encoder: Scalable Knowledge Distillation from Diffusion Models Paper • 2507.07104 • Published Jul 9, 2025 • 47
Play to Generalize: Learning to Reason Through Game Play Paper • 2506.08011 • Published Jun 9, 2025 • 15
Medical World Model: Generative Simulation of Tumor Evolution for Treatment Planning Paper • 2506.02327 • Published Jun 2, 2025 • 20
VideoAuteur: Towards Long Narrative Video Generation Paper • 2501.06173 • Published Jan 10, 2025 • 31