Reflect, Revise, Reuse: Training-Free Skill Evolution for GUI Agents Paper • 2609.17653 • Published 13 days ago • 45
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 21 days ago • 374
Beyond Solver Verdicts: Generative Reward Models for Autoformalization Paper • 2609.11085 • Published 18 days ago • 34
Scaling Automatic Research Agents via World Models Paper • 2608.12564 • Published about 1 month ago • 481
World-to-Wrist: Task-Conditioned Future Wrist Modeling for Fine-Grained Robot Manipulation Paper • 2608.05369 • Published Aug 5 • 27
Distill Where You Fail: Recovering Learning Signals of Negative RL-Groups from Adaptive Teacher Guidance Paper • 2608.00782 • Published Aug 1 • 18
CAPEval: A Decoupled Caption Evaluation across Understanding and Generation Paper • 2608.02589 • Published Aug 3 • 26
SKT: Skill-Use Training at Scale via Verified Synthetic Data Generation Paper • 2608.02287 • Published Aug 3 • 32