Foundations of Proactive Agents: Principles, Technical Layers, and Proactivity-Gym Paper • 2609.37267 • Published 9 days ago • 28
Learning Global Motion with Compact Gaussians for Feed-Forward 4D Reconstruction Paper • 2605.31595 • Published May 29
Imagine3D-LLM: Teaching MLLMs to Imagine 3D Scenes Before Answering Paper • 2609.38177 • Published 9 days ago • 72
EgoTools: Towards Tool-Centric Reasoning in Real-World Egocentric Videos Paper • 2609.39378 • Published 8 days ago • 67
World Observer: Joint Actor-Observer Generation for Persistent World Modeling Paper • 2610.02162 • Published 7 days ago • 85
Mid-Harness: Scaling Actions Between Model and Harness for Terminal Agents Paper • 2609.39982 • Published 8 days ago • 117
Imagine3D-LLM: Teaching MLLMs to Imagine 3D Scenes Before Answering Paper • 2609.38177 • Published 9 days ago • 72
MVTrack4Gen: Multi-View Point Tracking as Geometric Supervision for 4D Video Generation Paper • 2606.26087 • Published Jun 24 • 37
Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients Paper • 2606.18216 • Published Jun 16 • 65
MaskingDepth: Masked Consistency Regularization for Semi-supervised Monocular Depth Estimation Paper • 2212.10806 • Published Dec 21, 2022
Visual Representation Alignment for Multimodal Large Language Models Paper • 2509.07979 • Published Sep 9, 2025 • 84
TrackCraft3R: Repurposing Video Diffusion Transformers for Dense 3D Tracking Paper • 2605.12587 • Published May 12 • 37