Σ-Mem: An Online Reliability Memory for LLM-based Multi-Agent Systems Paper • 2607.27958 • Published 2 days ago • 8
ShadowDancer: Teaching Video World Models Any Action by Learning Unified Dynamics Representations from a Video and Its Shadow Paper • 2607.28362 • Published 2 days ago • 16
Echoverse: Deep, Evolving Environments for Training Computer-Use Agents at Scale Paper • 2607.28074 • Published 2 days ago • 9
INTACT: Isomorphic Intent-to-Action Learning for Search-Free World Models Paper • 2607.26056 • Published 4 days ago • 12
Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Paper • 2607.27919 • Published 2 days ago • 48
ACE-Data-0: Human-Centric Ambient Capture as Embodied Data Engine Paper • 2607.28625 • Published 2 days ago • 34
RefCaptioner: Multi-Reference Image-Grounded Video Captioning Paper • 2607.28509 • Published 2 days ago • 24
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published 2 days ago • 271
MPIE-Bench: Benchmarking Anatomically Plausible Multi-Person Interaction Editing Paper • 2607.27616 • Published 2 days ago • 36
Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Evaluation Paper • 2607.27816 • Published 2 days ago • 30
BM25 Wins at Scale: A Scaling Study of Retrieval-Augmented Generation Paradigms Paper • 2607.26497 • Published 2 days ago • 39
Beacon: Knowing When and How to Perform Agentic Visual Reasoning Paper • 2607.28595 • Published 2 days ago • 45
PhiZero: A World Model Built Around Physical Language Paper • 2607.28624 • Published 2 days ago • 151
Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering Paper • 2607.28568 • Published 2 days ago • 158
SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them Paper • 2607.27703 • Published 2 days ago • 21
Can AI agents conduct open-ended AI research? Early evidence from two case studies Paper • 2607.27191 • Published 3 days ago • 14
OmegaUse-OfficeVal: Benchmarking LLM Agents on Long-Horizon Office-Suite Tasks with Economic Grounding Paper • 2607.27155 • Published 3 days ago • 8