Model specification · NVIDIA
Nemotron 3 Nano 30B A3B
nvidia/nemotron-3-nano-30b-a3b
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
SpecificationsReported values
- Context window
- 262K
- Maximum output
- 262K
- Architecture
- text->text
- Tokenizer
- Other
- Knowledge cutoff
- —
- Moderated
- No
- Input modalities
- text
- Output modalities
- text
Full pricingUSD where reported
| Price | Standard | free |
|---|---|---|
| Input | $0.050 / M | free / M |
| Output | $0.200 / M | free / M |
Benchmark indexesReported indexes
- Intelligence
- 14.2
- Coding
- 14.4
- Agentic
- 2
- Design Arena Elo
- —
Example workload costsStandard pricing only
| Workload | Input / output | Cache share | Estimated cost |
|---|---|---|---|
| Quick check | 10K / 1K × 1 | 0% | $0.00070 |
| Long analysis | 100K / 4K × 1 | 50% | $0.00480 |
| Workday | 20K / 2K × 1,000 | 30% | $1.28 |
More from NVIDIAAll NVIDIA models
| Model | Context | Input / M | Output / M | Intel. |
|---|---|---|---|---|
| Nemotron 3 Ultra | 512K | $0.600 | $3.60 | 37.8 |
| Nemotron 3 Super | 1M | $0.085 | $0.400 | 25.4 |