RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 9 days ago • 216
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 16 days ago • 248
Where to Look Matters: On-Policy Self-Distillation for Long-Video Understanding Paper • 2608.25356 • Published Aug 26 • 20
Rubrics as Visual-Repair Context for Self-Evolving UI-to-Code Generation Paper • 2608.24138 • Published Aug 25 • 13
LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling Paper • 2605.08083 • Published May 8 • 71
Explore Data Left Behind in Reinforcement Learning for Reasoning Language Models Paper • 2511.04800 • Published Nov 6, 2025 • 1