Running 62 Don't Train the Model, Evolve the Harness 🌿 62 Evolving an agent's harness, not its model, on Harvey's LAB
nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 Text Generation • 32B • Updated 27 days ago • 995k • • 812
Running 226 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 226 Building and scaling RL environments for LLM training