Arm-wise Compositional Generalization in Dual-Arm Vision-Language-Action Models Paper • 2610.06184 • Published 4 days ago • 4
Learning Foresight without Explicit Trajectories for 3D Diffusion Policies Paper • 2609.20669 • Published 22 days ago • 9
CARE: Experience-Guided Atomic Corrective Execution for Vision-Language-Action Policies Paper • 2609.24118 • Published 18 days ago • 28
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 18 days ago • 223
JEPA-Anything: Learning Predictive Models across Different Worlds Paper • 2609.20800 • Published 22 days ago • 77
VABench: Measuring Embodied Spatial Intelligence through Visual Demonstrations, Active Perception, and Metric Control Paper • 2609.19554 • Published 22 days ago • 43
Attention-DP3: Spatially Object-aware 3D Diffusion Policy via Geometry-aligned Attentional Conditioning Paper • 2609.13318 • Published 29 days ago • 7
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation Paper • 2609.11115 • Published 29 days ago • 174
MA-VLA: Multi-Arm Vision-Language-Action Model for Collaboration and Compositional Generalization Paper • 2608.25864 • Published Aug 26 • 9
Super Star: Towards Streaming Real-time Interactive Agents for Digital Humans Paper • 2608.24909 • Published Jul 22 • 6
HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Paper • 2607.25895 • Published Jul 28 • 95
N_0-VTLA: Scaling Vision-Tactile-Language-Action Model with Latent Tactile Tokens Paper • 2607.23782 • Published Jul 26 • 80
N_0-TWAM: Scaling Tactile-Native World-Action Model for Contact-Rich Manipulation Paper • 2607.23783 • Published Jul 26 • 48
Skill-3D: Evolving Scene-Aware Skills for Agentic 3D Spatial Reasoning Paper • 2606.07436 • Published Jun 5 • 27
ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research Paper • 2606.07591 • Published May 28 • 106
LAP: Language-Action Pre-Training Enables Zero-shot Cross-Embodiment Transfer Paper • 2602.10556 • Published Feb 11 • 2
DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation Paper • 2601.22153 • Published Jan 29 • 77
AgentDoG: A Diagnostic Guardrail Framework for AI Agent Safety and Security Paper • 2601.18491 • Published Jan 26 • 127