ExpVoyager: Direct Experience Navigation for Dynamic Agent Skill Synthesis Paper • 2609.32630 • Published 6 days ago • 14
Skill2Env: Capability-Oriented Environment Synthesis from Skills for General Agents Paper • 2609.33772 • Published 5 days ago • 31
TraceDance: An Automated System for Building Agent Behavior Benchmarks from Real-World Agent Deployment Traces Paper • 2609.33295 • Published 5 days ago • 66
VisionHOPE: Visual Backbones as Self-Modifying Learning Systems Paper • 2609.33325 • Published 5 days ago • 304
IndicBankBench: Evaluating Safety and Reliability of Language Model Assistants in Indian Retail Banking Paper • 2609.29167 • Published 8 days ago • 21
Enhancing Photogrammetric Digital Surface Models with Pretrained Diffusion Models and Multimodal Conditioning Paper • 2609.31199 • Published 7 days ago • 14
PUBG Ally: A Conversational Embodied Agent as an AI Teammate Paper • 2609.29837 • Published 8 days ago • 25
Agent-Editing World Model: Rethinking World Modeling for LLM Agents Paper • 2609.28416 • Published 9 days ago • 42
IterSynth: Rethinking Deep Search Agents via Role-Decoupled Iterative Synthesis Paper • 2609.29444 • Published 8 days ago • 20
Qwen-Planner-Agent: A Closed-Loop AI-for-AI Framework for Real-World Mobile Planner Agents Paper • 2609.29892 • Published 8 days ago • 32
StudentSim: Training LLM-based Student Simulators Paper • 2609.01591 • Published about 1 month ago • 494
EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading Agents Paper • 2609.17632 • Published 17 days ago • 46
StableVQ: Practical Guidelines for Stable Vector-Quantized Tokenizer Training Paper • 2609.26774 • Published 10 days ago • 56
Emergent Collusion in Long-Horizon LLM Agent Interaction Paper • 2609.24967 • Published 11 days ago • 19
The Tasteful Agent: Measuring and Improving Taste in Long-Horizon Tasks Paper • 2609.25804 • Published 10 days ago • 161
LatentPort: Beyond KV Cache - Cross-Model Transfer of Recurrent Memory in Hybrid Language Models: A 4B-to-9B Hybrid-State Handoff Without Target Prefix Replay Paper • 2609.25053 • Published 25 days ago • 17
Harness-Zero: Harness Distillation via Agent-as-Harness Paper • 2609.24974 • Published 11 days ago • 37
EDGEGEN: Improving Tool-Calling Agents Beyond Happy Paths with Synthetic Edge Case Generation Paper • 2609.24115 • Published 11 days ago • 5