Improving Proactive AI Assistance with Hierarchical Procedural Understanding Paper • 2610.06505 • Published 4 days ago • 9
Attacca: Goal-Directed Control under State Continuity for Long-Horizon Embodied Agents Paper • 2610.07785 • Published 4 days ago • 5
MetaRubric: Learning to Reward for Rubric-Based Reinforcement Learning Paper • 2610.02824 • Published 8 days ago • 33
PaperCompiler: Faithful Paper-to-Code Generation via Repository-Level Specification Compilation Paper • 2609.02272 • Published Sep 2 • 7
MuseBench: Benchmarking Intent-Level Audiovisual Arts Understanding in MLLMs Paper • 2606.30026 • Published Jun 29 • 6
Confidence-Aware Tool Orchestration for Robust Video Understanding Paper • 2606.26904 • Published Jun 25 • 12
PhyMotion: Structured 3D Motion Reward for Physics-Grounded Human Video Generation Paper • 2605.14269 • Published May 14 • 7
Can Large Language Models Keep Up? Benchmarking Online Adaptation to Continual Knowledge Streams Paper • 2603.07392 • Published Mar 8 • 18
MA-EgoQA: Question Answering over Egocentric Videos from Multiple Embodied Agents Paper • 2603.09827 • Published Mar 10 • 30
AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories Paper • 2602.14941 • Published Feb 16 • 6
When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning Paper • 2602.08236 • Published Feb 9 • 9
Reliable and Responsible Foundation Models: A Comprehensive Survey Paper • 2602.08145 • Published Feb 4 • 8
Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Conversation Paper • 2601.00664 • Published Jan 2 • 58
WorldMM: Dynamic Multimodal Memory Agent for Long Video Reasoning Paper • 2512.02425 • Published Dec 2, 2025 • 25
Video-RTS: Rethinking Reinforcement Learning and Test-Time Scaling for Efficient and Enhanced Video Reasoning Paper • 2507.06485 • Published Jul 9, 2025 • 5