TERRA: Terrain-Aware Reconstruction, Retargeting and Control for Musculoskeletal Locomotion Paper • 2609.38653 • Published 3 days ago • 25
Running 250 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 250 Building and scaling RL environments for LLM training
No Prompt Left Behind: Exploiting Zero-Variance Prompts in LLM Reinforcement Learning via Entropy-Guided Advantage Shaping Paper • 2509.21880 • Published Sep 26, 2025 • 54