nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 Text Generation • 32B • Updated 12 days ago • 931k • • 805
Running 3.96k The Ultra-Scale Playbook 🌌 3.96k The ultimate guide to training LLM on large GPU Clusters