Learning to Move Before Learning to Do: Task-Agnostic pretraining for VLAs Paper • 2607.02466 • Published 30 days ago • 11
Retrospective Harness Optimization: Improving LLM Agents via Self-Preference over Trajectory Rollouts Paper • 2606.05922 • Published Jun 4 • 70
Learning from the Self-future: On-policy Self-distillation for dLLMs Paper • 2606.18195 • Published Jun 16 • 77
FRAPPE: Full Input, Residual Output Autoencoding with Projection Pursuit Encoder Paper • 2605.28992 • Published May 27 • 7
Measuring the Depth of LLM Unlearning via Activation Patching Paper • 2605.24614 • Published May 23 • 8
NSF-SciFy: Mining the NSF Awards Database for Scientific Claims Paper • 2503.08600 • Published May 25 • 4
Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players Paper • 2605.28816 • Published May 27 • 433
EvalVerse: Pipeline-Aware and Expert-Calibrated Benchmarking for Professional Cinematic Video Generation Paper • 2605.23271 • Published May 22 • 82
CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence Paper • 2605.12882 • Published May 13 • 274
ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling Paper • 2603.25746 • Published Mar 26 • 155
SocialOmni: Benchmarking Audio-Visual Social Interactivity in Omni Models Paper • 2603.16859 • Published Mar 17 • 248