Can Computation from Earlier Problems Help LLMs Solve New Ones? Paper • 2609.39394 • Published 6 days ago • 8
Agent-Editing World Model: Rethinking World Modeling for LLM Agents Paper • 2609.28416 • Published 13 days ago • 43
TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning Paper • 2608.04007 • Published Aug 4 • 20
Towards Verifiable Multimodal Deep Research: A Multi-Agent Harness for Interleaved Report Generation Paper • 2605.29861 • Published May 28 • 11
Agentic Fusion of Large Atomic and Language Models to Accelerate Superconductors Discovery Paper • 2604.23758 • Published Apr 29 • 8
MultiWorld: Scalable Multi-Agent Multi-View Video World Models Paper • 2604.18564 • Published Apr 20 • 49
MultiWorld: Scalable Multi-Agent Multi-View Video World Models Paper • 2604.18564 • Published Apr 20 • 49
ChemVTS-Bench: Evaluating Visual-Textual-Symbolic Reasoning of Multimodal Large Language Models in Chemistry Paper • 2511.17909 • Published Nov 22, 2025 • 1
AgentProcessBench: Diagnosing Step-Level Process Quality in Tool-Using Agents Paper • 2603.14465 • Published Mar 15 • 23
Budget-Constrained Agentic Large Language Models: Intention-Based Planning for Costly Tool Use Paper • 2602.11541 • Published Feb 12 • 2
LawThinker: A Deep Research Legal Agent in Dynamic Environments Paper • 2602.12056 • Published Feb 12 • 36
Budget-Constrained Agentic Large Language Models: Intention-Based Planning for Costly Tool Use Paper • 2602.11541 • Published Feb 12 • 2
Controllable Preference Optimization: Toward Controllable Multi-Objective Alignment Paper • 2402.19085 • Published Feb 29, 2024
OptDist: Learning Optimal Distribution for Customer Lifetime Value Prediction Paper • 2408.08585 • Published Aug 16, 2024