nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 Text Generation • 32B • Updated 24 days ago • 705k • • 822
Running 4.04k The Ultra-Scale Playbook 🌌 4.04k The ultimate guide to training LLM on large GPU Clusters