MiVLA: Towards Generalizable Vision-Language-Action Model with Human-Robot Mutual Imitation Pre-training Paper • 2512.15411 • Published Dec 17, 2025
InternVLA-A1.5: Unifying Understanding, Latent Foresight, and Action for Compositional Generalization Paper • 2607.04988 • Published 28 days ago • 28
WSA$_1$: a 3D-Centric World-Spatial-Action Model for Generalizable Robot Control Paper • 2607.03941 • Published about 1 month ago • 1