Evidence Attribution in Visual Document Understanding without Coordinates or Region Labels Paper • 2607.24651 • Published 5 days ago • 3
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources Paper • 2606.29538 • Published 16 days ago • 142
SearchOS-V1: Towards Robust Open-Domain Information-Seeking Agent Collaboration Paper • 2607.15257 • Published 16 days ago • 71
Bridging Interleaved Multi-Modal Reasoning as a Unified Decision Process Paper • 2607.03748 • Published 28 days ago • 41
sentence-transformers/paraphrase-multilingual-mpnet-base-v2 Sentence Similarity • 0.3B • Updated Aug 19, 2025 • 11.2M • • 484
Managing Procedural Memory in LLM Agents: Control, Adaptation, and Evaluation Paper • 2606.23127 • Published Jun 22 • 25
One Forward Beats Two: InnerZoom for Accurate and Efficient GUI Grounding Paper • 2606.30084 • Published Jun 29 • 8