Distilling Routed 3D Privilege for Spatial Reasoning in Vision-Language Models Paper • 2610.12355 • Published 3 days ago • 8
ViSkill: Reinforcing VLM Agents with Evolving Visual-Native Skills Paper • 2610.12403 • Published 3 days ago • 14
SpaceCast-Bench: Evaluating Predictive Spatial Reasoning in Vision-Language Models Paper • 2610.12402 • Published 3 days ago • 11
ComputerSD: Online Self-Distillation from Real-Time Feedback for Computer-Use Agents Paper • 2609.40253 • Published 10 days ago • 8
RetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning Paper • 2609.20784 • Published 24 days ago • 57
Reflect, Revise, Reuse: Training-Free Skill Evolution for GUI Agents Paper • 2609.17653 • Published 26 days ago • 45
PaperGym: Rubric-Centered Evolution for Research-Plan Generation Paper • 2608.31119 • Published Aug 31 • 33
Agent-G^2: Gaussian Guidance for Agentic Reinforcement Learning Paper • 2608.23318 • Published Aug 24 • 32
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning Paper • 2608.05987 • Published Aug 6 • 104
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Paper • 2607.26784 • Published Jul 29 • 29
ResearchStudio-Reel: Automate the Last Mile of Research from Paper to Poster, Video, and Blog Paper • 2607.04438 • Published Jul 5 • 62
AgenticDataBench: A Comprehensive Benchmark for Data Agents Paper • 2607.01647 • Published Jul 2 • 36
Perceive-to-Reason: Decoupling Perception and Reasoning for Fine-Grained Visual Reasoning Paper • 2607.01191 • Published Jul 1 • 18
minWM: A Full-Stack Open-Source Framework for Real-Time Interactive Video World Models Paper • 2605.30263 • Published May 28 • 58
view article Article Harness, Scaffold, and the AI Agent Terms Worth Getting Right sergiopaniego, ariG23498 • May 25 • 152
Enhancing Train-Free Infinite-Frame Generation for Consistent Long Videos Paper • 2605.18233 • Published May 18 • 92