Progress Reward Modeling for Robotic Learning: A Comprehensive Survey Paper • 2607.21655 • Published 12 days ago • 192
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU Paper • 2607.19191 • Published 13 days ago • 308
MuScriptor: An Open Model for Multi-Instrument Music Transcription Paper • 2607.08168 • Published 25 days ago • 21
Read It Back: Pretrained MLLMs Are Zero-Shot Reward Models for Text-to-Image Generation Paper • 2607.11886 • Published 21 days ago • 84
HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better Paper • 2607.04884 • Published 28 days ago • 9
RynnWorld-Teleop: An Action-Conditioned World Model for Digital Teleoperation Paper • 2607.06558 • Published 27 days ago • 79
Denser neq Better: Limits of On-Policy Self-Distillation for Continual Post-Training Paper • 2607.01763 • Published Jul 2 • 10
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published Jun 28 • 170
MemSlides: A Hierarchical Memory Driven Agent Framework for Personalized Slide Generation with Multi-turn Local Revision Paper • 2606.17162 • Published Jun 15 • 177
Escaping the Self-Confirmation Trap: An Execute-Distill-Verify Paradigm for Agentic Experience Learning Paper • 2606.24428 • Published Jun 23 • 52
DataClaw0: Agentic Tailoring Multimodal Data from Raw Streams Paper • 2606.21337 • Published Jun 19 • 75
Moebius: 0.2B Lightweight Image Inpainting Framework with 10B-Level Performance Paper • 2606.19195 • Published Jun 17 • 141