Milestones
- Nemotron 3 Nano released on Hugging Face15 Dec 2025Complete.
- Super and Ultra models releaseTarget Q2 2026Not yet reached.
Most important updates
Upcoming
- Q2 2026Super and Ultra models release (next)
by NVIDIAUS
Open-weight LLMs in three sizes using hybrid mixture-of-experts, activating a fraction of parameters per token. Designed for customization and multi-agent deployments.
Nano (30B, 3B active) released Dec 2025; Super (~100B, 10B active) and Ultra (~500B, 50B active) in H1 2026
Updated 15 Dec 2025Checked 10 Oct0 updates this week
Only a fraction of parameters activate per token. Routes different inputs to specialized sub-networks, reducing compute while handling diverse tasks.
NVIDIA released its training and safety datasets so enterprises can customize Nemotron on proprietary data without sharing with NVIDIA.
Designed for scenarios where multiple models work together, coordinating through text. Smaller active parameter count keeps latency low.
| Spec | Nemotron 3 |
|---|---|
| Nemotron 3 Nano size | 30B (3B active)R (reported) |
| Nemotron 3 Super size | ~100B (10B active)R (reported) |
| Nemotron 3 Ultra size | ~500B (50B active)R (reported) |
| Training datasets released | 3 trillion tokensR (reported) |
R reported by the company
NVIDIA
AI GPUs, rack systems, CUDA, networking
NVIDIA designs the GPUs, networking and software behind most AI training and inference. Its CUDA software made GPUs the default for AI. In October 2025 it became the first company worth $5 trillion.