Model specification · NVIDIA

Nemotron 3 Nano 30B A3B

nvidia/nemotron-3-nano-30b-a3b

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

SpecificationsReported values
Context window
262K
Maximum output
262K
Architecture
text->text
Tokenizer
Other
Knowledge cutoff
Moderated
No
Input modalities
text
Output modalities
text
Full pricingUSD where reported
PriceStandardfree
Input$0.050 / Mfree / M
Output$0.200 / Mfree / M
Benchmark indexesReported indexes
Intelligence
14.2
Coding
14.4
Agentic
2
Design Arena Elo
Example workload costsStandard pricing only
WorkloadInput / outputCache shareEstimated cost
Quick check10K / 1K × 10%$0.00070
Long analysis100K / 4K × 150%$0.00480
Workday20K / 2K × 1,00030%$1.28
More from NVIDIAAll NVIDIA models
ModelContextInput / MOutput / MIntel.
Nemotron 3 Ultra512K$0.600$3.6037.8
Nemotron 3 Super1M$0.085$0.40025.4