NVIDIA
Nemotron 3 Nano 30B A3B
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems.
At a glance
- Input / 1M
- $0.06
- Output / 1M
- $0.24
- Context
- 262K
- Max output
- 236K
- Intelligence
- 8.9
- Released
- Dec 2025
Pricing
Nemotron 3 Nano 30B A3B API pricing
Per-token rates, plus what common tasks actually cost.
| Input tokensPer 1M tokens | $0.06 |
|---|---|
| Output tokensPer 1M tokens | $0.24 |
| Cached input (read)Per 1M tokens | Not available |
| Cache writePer 1M tokens | Not available |
Chat message
2K in, 500 out
$0.24
per 1,000 requests
Coding task
30K in, 4K out
$2.76
per 1,000 requests
Long document summary
150K in, 2K out
$9.48
per 1,000 requests
Benchmarks
Nemotron 3 Nano 30B A3B benchmark scores
Independent scores from Artificial Analysis. Higher is better.
Intelligence index
8.9
Coding index
14.4
Agentic index
1
Specs
Context window and capabilities
What it can read, what it can write, and which API features it supports.
- Context window
- 262K tokens
- Max output
- 236K tokens
- Input types
- Text
- Output types
- Text
- Reasoning effort
- Not available
- Release date
- December 14, 2025
- Tool calling
- Structured outputs
- JSON mode
- Reasoning
- Temperature
- Stop sequences
- Deterministic seed
- Verbosity control
Compare
Nemotron 3 Nano 30B A3B vs other models
See how Nemotron 3 Nano 30B A3B stacks up head to head on price, benchmarks, and features.
Nemotron 3 Nano 30B A3B vs Gemini 2.5 Flash Lite
NVIDIA vs Google
See comparisonNemotron 3 Nano 30B A3B vs gpt-oss-120b
NVIDIA vs OpenAI
See comparisonNemotron 3 Nano 30B A3B vs Nemotron 3.5 Lightning
NVIDIA vs NVIDIA
See comparisonNemotron 3 Nano 30B A3B vs Nemotron 3 Super
NVIDIA vs NVIDIA
See comparisonNemotron 3 Nano 30B A3B vs Qwen3 30B A3B Instruct 2507
NVIDIA vs Qwen
See comparisonNemotron 3 Nano 30B A3B vs Granite 4.2 8B
NVIDIA vs IBM
See comparisonAlternatives
Nemotron 3 Nano 30B A3B alternatives
Similar models worth considering, and where each one has the edge.
NVIDIA
Nemotron 3.5 Lightning
OpenAI
gpt-oss-120b
IBM
Granite 4.2 8B
Want to learn about other models? Browse all models
FAQ
Frequently asked questions
How much does Nemotron 3 Nano 30B A3B cost?
Nemotron 3 Nano 30B A3B costs $0.06 per 1M input tokens and $0.24 per 1M output tokens through the API. A typical chat message costs about $0.24 per 1,000 messages.
What is the context window of Nemotron 3 Nano 30B A3B?
Nemotron 3 Nano 30B A3B accepts up to 262K tokens of input and can write up to 236K tokens in a single response.
What is the best alternative to Nemotron 3 Nano 30B A3B?
Nemotron 3.5 Lightning from NVIDIA is a strong alternative. It offers: 1.8x cheaper output, higher intelligence score, newer release.
Can I use Nemotron 3 Nano 30B A3B alongside other models?
Yes. Shortcut Chat sends one prompt to Nemotron 3 Nano 30B A3B and any other models you pick at the same time, so you can compare answers side by side.
Try Nemotron 3 Nano 30B A3B next to every other model.
Use Nemotron 3 Nano 30B A3B alongside all the other models. Use one at a time, or multiple together. Get the best answer from every AI.