Agensh: Scaling Organizational Intelligence to 1,024 Agents Paper • 2609.26781 • Published 6 days ago • 27
The Tasteful Agent: Measuring and Improving Taste in Long-Horizon Tasks Paper • 2609.25804 • Published 6 days ago • 159
MemoryAthena: Adaptive Routing over Latent and Generated Memories Paper • 2609.25853 • Published 6 days ago • 6
MemBodied: Recurrent Associative Memory for Vision-Language-Action Models Paper • 2609.28256 • Published 5 days ago • 13
Just-in-Time Memory: Learning to Curate Task-Adaptive Memory for LLM Agents Paper • 2609.27334 • Published 5 days ago • 41
The Past Frames the Future: Memory for Autoregressive Video Generation Paper • 2609.28466 • Published 5 days ago • 50
SpeakerMem-R1: Speaker-Centered Dual-Track Memory for Multi-Party Dialogue Paper • 2609.26780 • Published 6 days ago • 90
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation Paper • 2609.11115 • Published 18 days ago • 172
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 14 days ago • 247
Memory as Plans: World-Action Modeling with Memory-Grounded Planning Paper • 2609.11561 • Published 18 days ago • 41
LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks Paper • 2608.01964 • Published Aug 3 • 188
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement Paper • 2607.23802 • Published Jul 26 • 96
Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Paper • 2607.27919 • Published Jul 30 • 62
Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Paper • 2607.22529 • Published Jul 24 • 48
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published Jul 14 • 126