Annotations as Rollouts: Efficient and Scalable Reinforcement Learning for Video MLLMs Paper • 2608.20492 • Published Aug 20 • 87
Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments Paper • 2605.30280 • Published May 28 • 146
SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer Paper • 2605.15178 • Published May 14 • 92
RobustFlow: Towards Robust Agentic Workflow Generation Paper • 2509.21834 • Published Sep 26, 2025 • 2
RemoteZero: Geospatial Reasoning with Zero Human Annotations Paper • 2605.04451 • Published May 6 • 7
RemoteAgent: Bridging Vague Human Intents and Earth Observation with RL-based Agentic MLLMs Paper • 2604.07765 • Published Apr 12