Weak-to-Strong Generalization via Direct On-Policy Distillation Paper • 2607.05394 • Published 21 days ago • 138
Perceptual Flow Matching for Few-Step Generative Modeling Paper • 2607.03524 • Published 26 days ago • 18
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published about 1 month ago • 170
awrenn53/gr00t-n17-libero-goal-oss-h16-adamw-gbs640-s42-head9adea3f9-012000-20260706 3B • Updated 22 days ago • 13 • 1
UI-KOBE: Knowledge-Oriented Behavior Exploration for Lightweight Graph-Guided GUI Agents Paper • 2605.29534 • Published May 28 • 15
AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Question Answering? Paper • 2605.28255 • Published May 27 • 1
SAM: State-Adaptive Memory for Long-Horizon Reasoning Agent Paper • 2605.24468 • Published May 23 • 9
Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players Paper • 2605.28816 • Published May 27 • 433
Segment Anything with Motion, Geometry, and Semantic Adaptation for Complex Nonlinear Visual Object Tracking Paper • 2605.22538 • Published May 21 • 6