True2456/nemotron-3-super-120b-cybersecurity-theory-lora-mlx Text Generation • Updated 14 days ago • 2
Principled Analysis of Deep Reinforcement Learning Evaluation and Design Paradigms Paper • 2607.07769 • Published 22 days ago • 10
Running Featured 392 Bonsai 27B WebGPU Kernels 🌳 392 Run a 1-bit 27B LLM locally in your browser on WebGPU
nvidia/parakeet-tdt-0.6b-v3 Automatic Speech Recognition • 0.6B • Updated about 1 month ago • 150k • • 1.02k
Running 203 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 203 Building and scaling RL environments for LLM training