UniEvo-VL: An On-policy Self-Distillation Training Recipe for Multimodal Model Self-improvement Paper • 2609.38721 • Published 7 days ago • 292
Transferring the Intelligence of VLMs to Robotic Control Paper • 2609.22966 • Published 18 days ago • 120
NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness Paper • 2609.08183 • Published 29 days ago • 328