Echoverse: Deep, Evolving Environments for Training Computer-Use Agents at Scale Paper • 2607.28074 • Published 5 days ago • 10
LLM-as-a-Coach: Experiential Learning for Non-Verifiable Tasks Paper • 2607.18110 • Published 15 days ago • 14
Object-Centric Residual RL for Zero-Shot Sim-to-Real VLA Enhancement Paper • 2606.18953 • Published Jun 17 • 7
ReVision: Scaling Computer-Use Agents via Temporal Visual Redundancy Reduction Paper • 2605.11212 • Published Jun 5 • 5
Rebellious Student: Reversing Teacher Signals for Reasoning Exploration with Self-Distilled RLVR Paper • 2605.10781 • Published May 11 • 17
World-R1: Reinforcing 3D Constraints for Text-to-Video Generation Paper • 2604.24764 • Published Apr 27 • 119
Mind's Eye: A Benchmark of Visual Abstraction, Transformation and Composition for Multimodal LLMs Paper • 2604.16054 • Published Apr 17 • 1
BizGenEval: A Systematic Benchmark for Commercial Visual Content Generation Paper • 2603.25732 • Published Mar 26 • 11
Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs? Paper • 2603.24472 • Published Mar 25 • 57
Understanding Reasoning in LLMs through Strategic Information Allocation under Uncertainty Paper • 2603.15500 • Published Mar 16 • 12
Scaling Data Difficulty: Improving Coding Models via Reinforcement Learning on Fresh and Challenging Problems Paper • 2603.07779 • Published Mar 8 • 5
Breaking Training Bottlenecks: Effective and Stable Reinforcement Learning for Coding Models Paper • 2603.07777 • Published Mar 8 • 5
Sparse-BitNet: 1.58-bit LLMs are Naturally Friendly to Semi-Structured Sparsity Paper • 2603.05168 • Published Mar 5 • 6