Omni-Decision: Evidence-Ledger Planning for Omni-Modal Agents Paper • 2607.11433 • Published 14 days ago • 30
What Makes World Action Models Generalize? An Empirical Study of Test-Time Future Modeling Paper • 2609.34981 • Published 9 days ago • 136
Morphometric Imitation: From Morphology and Contact Aware Hand Retargeting to Sim-to-Real Visuomotor Policy Paper • 2609.28660 • Published 15 days ago • 16
VLA-Precision: Asymmetric Co-Bootstrapping for Efficient Real-World Online RL of Vision-Language-Action Models Paper • 2609.04355 • Published 20 days ago • 9
MemBodied: Recurrent Associative Memory for Vision-Language-Action Models Paper • 2609.28256 • Published 15 days ago • 17
Towards Full Pipeline FP8 Reinforcement Learning for LLMs Paper • 2609.22870 • Published 19 days ago • 18
One to More, More to One: Category-Aware Iterative Expert Training for Software Engineering Agents Paper • 2609.23377 • Published 18 days ago • 50
Jev-Mem: System-One-Controlled Agentic Memory for Efficient AI Agents Paper • 2609.23986 • Published 17 days ago • 28
From Pretraining to Proficiency: Real-World Subtask RL for Long-Horizon Manipulation with Minimal Human Intervention Paper • 2609.21788 • Published 20 days ago • 13
MintAct: A Unified Visual Agent for Digital Environments Paper • 2609.22083 • Published 20 days ago • 35
OmniVChat: Synthesizing, Benchmarking, and Training for Native Audio-Visual Dialogue Paper • 2609.21465 • Published 20 days ago • 151
BI-Agent and BI-Bench: Towards Automating End-to-End Business Intelligence Paper • 2609.20886 • Published 22 days ago • 30
CodeMidas: Scaling Agentic Coding RL Environments from Code Itself Paper • 2609.22068 • Published 20 days ago • 138
Region-Level Policy Optimization for Fine-grained MLLM Perception Paper • 2609.19745 • Published 21 days ago • 41
RetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning Paper • 2609.20784 • Published 21 days ago • 57
Don't Mask the Environment: Observation Supervision Changes How Agents Explore Under RL Paper • 2609.20715 • Published 21 days ago • 44
SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness Paper • 2609.20519 • Published 21 days ago • 139
Fingers as Legs: Learning Self-Supported Locomotion and Manipulation with an Anthropomorphic Hand Paper • 2609.17172 • Published 23 days ago • 5