Running 59 Don't Train the Model, Evolve the Harness 🌿 59 Evolving an agent's harness, not its model, on Harvey's LAB
nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 Text Generation • 32B • Updated 14 days ago • 931k • • 805
Running 204 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 204 Building and scaling RL environments for LLM training