SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness Paper • 2609.20519 • Published 9 days ago • 133
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 19 days ago • 373
Steering Geometry: Validating Human Value Geometry in LLM Steering Space Paper • 2609.06289 • Published 21 days ago • 31
WorldReward: Reward Modeling for Camera-Conditioned World Models Paper • 2609.03952 • Published 23 days ago • 27
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published about 1 month ago • 155
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 23 days ago • 186
Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills Paper • 2609.02749 • Published 24 days ago • 402
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 265
Learning from Failures: Retrieval-Centric CoT via Hard Negatives for Unified Multimodal Retrieval Paper • 2608.06060 • Published Aug 6 • 41