TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training Paper • 2607.05804 • Published 25 days ago • 19
fpadovani/dan-latn-10mb-after-ppt-shuff-dyck-10mb-ckpt500_seed10 Text Generation • 39.1M • Updated 25 days ago • 922 • 1
FRAPPE: Full Input, Residual Output Autoencoding with Projection Pursuit Encoder Paper • 2605.28992 • Published May 27 • 7
SimuWoB: Simulating Real-World Mobile Apps for Fast and Faithful GUI Agent Benchmarking Paper • 2605.25160 • Published May 24 • 9
AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration Paper • 2605.20025 • Published May 19 • 191
Distilling Long-CoT Reasoning through Collaborative Step-wise Multi-Teacher Decoding Paper • 2605.02290 • Published May 4 • 42