← Models
Nemotron 3 Nano (30B A3B)
NVIDIA · open weight · nemotron-3-nano-30b-a3b
Nemotron 3 Nano is a 31.6B hybrid MoE model optimized for fast, long‑context agentic reasoning. It mixes Mamba‑2 and Transformer layers with a sparse MoE router (~3.6B active params per token) to deliver up to 4× higher throughput than Nemotron 2 and strong accuracy across math, coding, and tools. It supports a 1M‑token context window, offers Reasoning ON/OFF and a thinking‑budget to control costs, and ships with open weights, data, and RL tooling (NeMo Gym/RL). Released Dec 15, 2025 under the NVIDIA Open Model License, it’s built as the efficient backbone for multi‑agent systems at scale.
Benchmark scores
| Benchmark | Score |
|---|---|
| AIME 2025 | 0.99 |
| MMLU-Pro | 0.78 |
| GPQA | 0.75 |
| SWE-Bench Verified | 0.39 |
| Humanity's Last Exam | 0.15 |
Pricing
- DeepInfra$0.05 / $0.20
Input / output per 1M tokens
AA metrics
No Artificial Analysis link yet.
Arena Elo
- No Arena snapshot linked.
