LUMOS: A Semantic Operating-System Layer for Accessibility-Grounded AI Agents Paper • 2606.30697 • Published Jun 29 • 7
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published Jun 28 • 170
BioTool: A Comprehensive Tool-Calling Dataset for Enhancing Biomedical Capabilities of Large Language Models Paper • 2605.05758 • Published May 7 • 5