4Director: Controlling Video World Models with Rigid 3D Geometry Paper • 2610.02160 • Published 1 day ago • 13
Beyond Memory: Harnessing Long-Horizon Agents with Explicit Belief States Paper • 2610.01415 • Published 1 day ago • 62
ActiveSaddler: Automated Curriculum Learning for Agent Harness Optimization Paper • 2610.00906 • Published 1 day ago • 34
AutoGUIWorld: Image Generators as Visual World Models for GUI Agent Paper • 2610.01215 • Published 1 day ago • 28
The Evolution of Attention in Large Language Models: Mechanisms, Trade-offs, and Emerging Trends Paper • 2609.39661 • Published 3 days ago • 6
Learning Meta-Skills for Agent Harness Design in Test-Time AI4AI Paper • 2609.38143 • Published 4 days ago • 78
PivotOPD: Learning to Recover from Pivotal Mistakes in Multi-Turn Agents Paper • 2609.40285 • Published 3 days ago • 20
Mid-Harness: Scaling Actions Between Model and Harness for Terminal Agents Paper • 2609.39982 • Published 3 days ago • 108
RSIGame: Autonomous Agentic Game Development with Recursive Self-improvement Paper • 2609.39045 • Published 3 days ago • 77
Systematically Exploring the Capabilities of GPT-6 Astra as Embodied Policies Paper • 2609.38537 • Published 4 days ago • 29
AREX-2: Advancing Self-Improving Agents through Long-Horizon Reflective Tasks Paper • 2609.38288 • Published 4 days ago • 126
StoryEngine: A State-Grounded Agentic Framework for Video Storytelling Paper • 2609.33627 • Published 6 days ago • 17
What Makes World Action Models Generalize? An Empirical Study of Test-Time Future Modeling Paper • 2609.34981 • Published 4 days ago • 118
Chinese-Jev: Bringing System One Model to Chinese-Language Tasks Paper • 2609.36965 • Published 4 days ago • 17