From Pretraining to Proficiency: Real-World Subtask RL for Long-Horizon Manipulation with Minimal Human Intervention Paper • 2609.21788 • Published 15 days ago • 13
BI-Agent and BI-Bench: Towards Automating End-to-End Business Intelligence Paper • 2609.20886 • Published 17 days ago • 30
OmniVChat: Synthesizing, Benchmarking, and Training for Native Audio-Visual Dialogue Paper • 2609.21465 • Published 15 days ago • 150
MintAct: A Unified Visual Agent for Digital Environments Paper • 2609.22083 • Published 15 days ago • 35
Don't Mask the Environment: Observation Supervision Changes How Agents Explore Under RL Paper • 2609.20715 • Published 16 days ago • 44
Region-Level Policy Optimization for Fine-grained MLLM Perception Paper • 2609.19745 • Published 16 days ago • 41
RetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning Paper • 2609.20784 • Published 16 days ago • 57