EgoTools: Towards Tool-Centric Reasoning in Real-World Egocentric Videos Paper • 2609.39378 • Published 3 days ago • 22
4Director: Controlling Video World Models with Rigid 3D Geometry Paper • 2610.02160 • Published 1 day ago • 13
Beyond Memory: Harnessing Long-Horizon Agents with Explicit Belief States Paper • 2610.01415 • Published 1 day ago • 67
ActiveSaddler: Automated Curriculum Learning for Agent Harness Optimization Paper • 2610.00906 • Published 1 day ago • 35
AutoGUIWorld: Image Generators as Visual World Models for GUI Agent Paper • 2610.01215 • Published 1 day ago • 31
The Evolution of Attention in Large Language Models: Mechanisms, Trade-offs, and Emerging Trends Paper • 2609.39661 • Published 3 days ago • 7
Learning Meta-Skills for Agent Harness Design in Test-Time AI4AI Paper • 2609.38143 • Published 4 days ago • 79
PivotOPD: Learning to Recover from Pivotal Mistakes in Multi-Turn Agents Paper • 2609.40285 • Published 3 days ago • 21
Mid-Harness: Scaling Actions Between Model and Harness for Terminal Agents Paper • 2609.39982 • Published 3 days ago • 109
RSIGame: Autonomous Agentic Game Development with Recursive Self-improvement Paper • 2609.39045 • Published 3 days ago • 78
Systematically Exploring the Capabilities of GPT-6 Astra as Embodied Policies Paper • 2609.38537 • Published 4 days ago • 30
AREX-2: Advancing Self-Improving Agents through Long-Horizon Reflective Tasks Paper • 2609.38288 • Published 4 days ago • 127
StoryEngine: A State-Grounded Agentic Framework for Video Storytelling Paper • 2609.33627 • Published 6 days ago • 17
What Makes World Action Models Generalize? An Empirical Study of Test-Time Future Modeling Paper • 2609.34981 • Published 4 days ago • 119