False Frontiers: Diagnosing and Mitigating Co-Cheating in Self-Evolving Search Agents Paper • 2609.39102 • Published 9 days ago • 676 • 3
False Frontiers: Diagnosing and Mitigating Co-Cheating in Self-Evolving Search Agents Paper • 2609.39102 • Published 9 days ago • 676
TwinRouterBench: Fast Static and Live Dynamic Evaluation for Realistic Agentic LLM Routing Paper • 2605.18859 • Published May 14 • 5
The Teacher Is a Direction, Not a Destination: Extrapolating RL-Induced Representation Residuals in On-Policy Distillation Paper • 2609.36484 • Published 10 days ago • 580
TermiGen: High-Fidelity Environment and Robust Trajectory Synthesis for Terminal Agents Paper • 2602.07274 • Published Feb 6 • 52
ECHO-2: A Large-Scale Distributed Rollout Framework for Cost-Efficient Reinforcement Learning Paper • 2602.02192 • Published Feb 2 • 13
Symphony: A Decentralized Multi-Agent Framework for Scalable Collective Intelligence Paper • 2508.20019 • Published Aug 27, 2025 • 2
SEDM: Scalable Self-Evolving Distributed Memory for Agents Paper • 2509.09498 • Published Sep 11, 2025 • 3
Lattica: A Decentralized Cross-NAT Communication Framework for Scalable AI Inference and Training Paper • 2510.00183 • Published Sep 30, 2025 • 8 • 1
Lattica: A Decentralized Cross-NAT Communication Framework for Scalable AI Inference and Training Paper • 2510.00183 • Published Sep 30, 2025 • 8