Weak-to-Strong Generalization via Direct On-Policy Distillation Paper • 2607.05394 • Published 23 days ago • 139
Precise Action-to-Video Generation Through Visual Action Prompts Paper • 2508.13104 • Published Aug 18, 2025 • 11
Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training Paper • 2501.11425 • Published Jan 20, 2025 • 108