AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 4 days ago • 140
NVIDIA-labs OO Agents: Native Python Object-Oriented Agents Paper • 2607.20709 • Published 5 days ago • 26
stefanocarrera/sqlautophagycode_D_test_Qwen3-8B_t1.25_g7_run0_metrics Viewer • Updated 7 days ago • 579 • 46 • 1
Navigating the Mirage: A Dual-Path Agentic Framework for Robust Misleading Chart Question Answering Paper • 2603.28583 • Published 13 days ago • 17
DuoMem: Towards Capable On-Device Memory Agents via Dual-Space Distillation Paper • 2606.29961 • Published 28 days ago • 10
Learning from Language Feedback via Variational Policy Distillation Paper • 2605.15113 • Published May 18 • 13
Anti-Self-Distillation for Reasoning RL via Pointwise Mutual Information Paper • 2605.11609 • Published May 12 • 196
Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization Paper • 2605.09996 • Published May 11 • 8