SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness Paper • 2609.20519 • Published 11 days ago • 137
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 21 days ago • 374
Scaling Automatic Research Agents via World Models Paper • 2608.12564 • Published about 1 month ago • 481
Steering Geometry: Validating Human Value Geometry in LLM Steering Space Paper • 2609.06289 • Published 23 days ago • 31
WorldReward: Reward Modeling for Camera-Conditioned World Models Paper • 2609.03952 • Published 25 days ago • 27
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published Aug 27 • 155
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 25 days ago • 186
Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills Paper • 2609.02749 • Published 26 days ago • 403
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 265
Learning from Failures: Retrieval-Centric CoT via Hard Negatives for Unified Multimodal Retrieval Paper • 2608.06060 • Published Aug 6 • 41
UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models Paper • 2608.04701 • Published Aug 5 • 9
StyleForge: Indoor Furniture Styling by Counterfactual Reasoning in a Hypergraph Field Paper • 2608.01954 • Published Aug 3 • 13