RetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning Paper • 2609.20784 • Published 14 days ago • 57
Scaling Creative Writing Beyond Story-Centric Data with Attribute-Guided Genre Expansion Paper • 2608.13947 • Published Aug 14 • 15