Learning from Runtime Feedback through Failure-Bank Self-Evolution for Vision-Language-Action Models Paper • 2609.39820 • Published 9 days ago • 29
What Does Privileged Information Add to On-Policy Self-Distillation? Paper • 2609.20612 • Published 22 days ago • 36
HypoEvolve: Genetic Algorithms Enable Multi-Agent LLMs to Discover Scientific Hypotheses Paper • 2609.15938 • Published 25 days ago • 31
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published Sep 7 • 376