Dream-Tac: A Unified Tactile World Action Model for Contact-Rich Robot Manipulation Paper • 2606.08737 • Published Jun 7
Vision-Language-Action Models for Autonomous Driving: Past, Present, and Future Paper • 2512.16760 • Published Dec 18, 2025 • 15
VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer Paper • 2512.11891 • Published Dec 9, 2025 • 10
RynnVLA-002: A Unified Vision-Language-Action and World Model Paper • 2511.17502 • Published Nov 21, 2025 • 28
RynnVLA-001: Using Human Demonstrations to Improve Robot Manipulation Paper • 2509.15212 • Published Sep 18, 2025 • 22
Towards Affordance-Aware Robotic Dexterous Grasping with Human-like Priors Paper • 2508.08896 • Published Aug 12, 2025 • 12
MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems Paper • 2503.16549 • Published Mar 19, 2025 • 15
LargeAD: Large-Scale Cross-Sensor Data Pretraining for Autonomous Driving Paper • 2501.04005 • Published Jan 7, 2025 • 1
iControl3D: An Interactive System for Controllable 3D Scene Generation Paper • 2408.01678 • Published Aug 3, 2024
Using Left and Right Brains Together: Towards Vision and Language Planning Paper • 2402.10534 • Published Feb 16, 2024 • 1
Segment Any Point Cloud Sequences by Distilling Vision Foundation Models Paper • 2306.09347 • Published Jun 15, 2023 • 1
JoyAI-VL-Interaction: Real-Time Vision-Language Interaction Intelligence Paper • 2606.14777 • Published Jun 10 • 216
Vision-Language-Action Models for Autonomous Driving: Past, Present, and Future Paper • 2512.16760 • Published Dec 18, 2025 • 15
VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer Paper • 2512.11891 • Published Dec 9, 2025 • 10
RynnVLA-002: A Unified Vision-Language-Action and World Model Paper • 2511.17502 • Published Nov 21, 2025 • 28