4DAnyone: Create Anyone in 4D from a Casual Monocular Video Paper • 2608.20335 • Published Aug 20 • 86
Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization Paper • 2608.26103 • Published Aug 26 • 26
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence Paper • 2607.07675 • Published Jul 8 • 64
robbyant/lingbot-world-v2-14b-causal-fast-diffusers Image-to-Video • 19B • Updated Jul 8 • 314 • 14
robbyant/lingbot-world-v2-14b-causal-fast-diffusers Image-to-Video • 19B • Updated Jul 8 • 314 • 14