HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Paper • 2607.25895 • Published 5 days ago • 146
A New Role for Relevance: Guiding Corpus Interaction in Agentic Search Paper • 2607.24223 • Published 6 days ago • 91
ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Paper • 2607.24743 • Published 6 days ago • 11
The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation Paper • 2607.24720 • Published 6 days ago • 25
K12-KGraph: A Curriculum-Aligned Knowledge Graph for Benchmarking and Training Educational LLMs Paper • 2605.09635 • Published 10 days ago • 63
Recurrent Sinusoidal INRs for Efficient High-Fidelity Representation Paper • 2607.21485 • Published 10 days ago • 9
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 10 days ago • 150
SeededGrasp: Language-Guided Grasping in Complex Scenes with Multiple Embodiments Paper • 2607.20207 • Published 11 days ago • 4
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs Paper • 2607.17423 • Published 14 days ago • 166
RAGU: A Multi-Step GraphRAG Engine with a Compact Domain-Adapted LLM Paper • 2607.11683 • Published 20 days ago • 148
KnowAct-GUIClaw: Know Deeply, Act Perfectly, Personal GUI Assistant with Self-Evolving Memory and Skill Paper • 2607.12625 • Published 18 days ago • 80
Navigating the Mirage: A Dual-Path Agentic Framework for Robust Misleading Chart Question Answering Paper • 2603.28583 • Published 19 days ago • 17
Automating the Design of Embodied Agent Architectures Paper • 2606.30111 • Published 30 days ago • 13
Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation Paper • 2607.05382 • Published 24 days ago • 87
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence Paper • 2607.07675 • Published 25 days ago • 64
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning Paper • 2606.31825 • Published Jun 30 • 29
Discrete Diffusion Language Models for Interactive Radiology Report Drafting Paper • 2607.01436 • Published Jul 1 • 11
SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe Paper • 2607.03451 • Published 30 days ago • 34