The information geometry of large language models is shared, learned, and controllable Paper • 2609.11063 • Published 19 days ago • 6
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 8 days ago • 211
OmniEdu: Open Foundation Models for Learning and Teaching Paper • 2609.23088 • Published 10 days ago • 234
Realtime-Venus: A full-duplex interaction system with asynchronous delegation Paper • 2609.13814 • Published 17 days ago • 224
WorldCrafter: Consistent Video World Model with Implicit 3D-aware Memory Paper • 2609.24984 • Published 8 days ago • 156
Zetta ζ: An Efficient Closed-Loop Embodied Harness for Self-Evolving Physical Intelligence Paper • 2608.16590 • Published Aug 17 • 153
Scaling Automatic Research Agents via World Models Paper • 2608.12564 • Published about 1 month ago • 482
Agora: Git as Shared Memory for Collective AutoResearch Paper • 2609.18094 • Published 13 days ago • 58
An Empirical Study of Harness Design for Coding Agents Paper • 2609.20804 • Published 12 days ago • 91
WorldSculpt: Generating Compositional Worlds from Grounded Videos Paper • 2609.05416 • Published 25 days ago • 26
Bilevel Coordinated Reflection: A Game-Theoretic Approach to Multi-Agent LLM Systems Paper • 2609.02750 • Published 27 days ago • 144
Terminal-Universe: Turning Agent Trajectories into Scalable Terminal Environments Paper • 2609.04148 • Published 26 days ago • 247
HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness? Paper • 2609.01437 • Published 28 days ago • 221
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement Paper • 2608.31046 • Published 29 days ago • 97
Understanding Evolution Strategies for LLM Reasoning: Broader Reasoning Coverage than GRPO Paper • 2608.27351 • Published Aug 27 • 22
Co-RL: Unsupervised Reasoning Emerges from Diverse Cohort in Multi-agent RL Paper • 2608.17253 • Published Aug 19 • 99