OpenForgeRL: Train Harness-native Agents in Any Environment Paper • 2607.21557 • Published 4 days ago • 7
OpenForgeRL: Train Harness-native Agents in Any Environment Paper • 2607.21557 • Published 4 days ago • 7
Synthetic Computers at Scale for Long-Horizon Productivity Simulation Paper • 2604.28181 • Published Apr 30 • 21
Synthetic Computers at Scale for Long-Horizon Productivity Simulation Paper • 2604.28181 • Published Apr 30 • 21
GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL Paper • 2602.22190 • Published Feb 25 • 17
Dyna-Mind: Learning to Simulate from Experience for Better AI Agents Paper • 2510.09577 • Published Oct 10, 2025 • 9
Phi-4-Mini-Reasoning: Exploring the Limits of Small Reasoning Language Models in Math Paper • 2504.21233 • Published Apr 30, 2025 • 50
Reinforcement Learning for Reasoning in Large Language Models with One Training Example Paper • 2504.20571 • Published Apr 29, 2025 • 99
Improving Autonomous AI Agents with Reflective Tree Search and Self-Learning Paper • 2410.02052 • Published Oct 2, 2024 • 9
Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing Paper • 2404.12253 • Published Apr 18, 2024 • 55
Teaching Language Models to Self-Improve through Interactive Demonstrations Paper • 2310.13522 • Published Oct 20, 2023 • 12