Progress Reward Modeling for Robotic Learning: A Comprehensive Survey Paper • 2607.21655 • Published 19 days ago • 193
Transcription Policy as a Latent Variable: Activating Controllable Verbatim ASR with Word-Level Timing Paper • 2607.18934 • Published 20 days ago • 6
KnowAct-GUIClaw: Know Deeply, Act Perfectly, Personal GUI Assistant with Self-Evolving Memory and Skill Paper • 2607.12625 • Published 26 days ago • 81
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU Paper • 2607.19191 • Published 19 days ago • 309
RynnBrain 1.1: Towards More Capable and Generalizable Embodied Foundation Model Paper • 2607.17977 • Published 21 days ago • 198
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs Paper • 2607.17423 • Published 22 days ago • 167
X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras Paper • 2607.12993 • Published 27 days ago • 131
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published 27 days ago • 232
ACID: Action Consistency via Inverse Dynamics for Planning with World Models Paper • 2607.02403 • Published Jul 2 • 23
MemSlides: A Hierarchical Memory Driven Agent Framework for Personalized Slide Generation with Multi-turn Local Revision Paper • 2606.17162 • Published Jun 15 • 177
OpenSTBench: Beyond Semantic Evaluation for Speech Translation Paper • 2605.30792 • Published May 29 • 3
Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Paper • 2605.24681 • Published May 23 • 5