NVIDIA: Nemotron 3 Nano 30B A3B
The NVIDIA Nemotron 3 Nano 30B A3B is designed for handling extensive text inputs with a context length of up to 262,144 tokens and can generate up to 228,000 output tokens. It supports tools and reasoning, making it suitable for complex tasks that require structured data processing. This model stands out with its comprehensive input modality focusing on text alone. For those prioritizing a blend of performance and cost-effectiveness, the Nemotron 3 Nano 30B A3B is well-worth considering. With a blended benchmark score of 10.7 from a single independent source and pricing at $0.05 per million input tokens and $0.2 per million output tokens, it offers competitive value for its capabilities. While the benchmark coverage is comprehensive, the model's specific use case should be evaluated against your needs to ensure optimal performance and cost efficiency.
Benchmark results
Independent, published benchmarks. Blended score 10.6 across 1 benchmark, last refreshed 2026-07-27. How scoring works →
| Benchmark | Measures | Score |
|---|---|---|
| AI Index | broad capability composite | 12.2 |
- Model ID
- nvidia/nemotron-3-nano-30b-a3b
- Vendor
- nvidia
- Released
- December 2025
- Tokenizer
- Other
- Input Modalities
- text
- Output Modalities
- text
- Max Output
- 228,000 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- ✓ supported
- Vision
- text only
- Audio
- no
- Moderated
- no
What it costs in practice
Computed from the current $0.05/M input and $0.20/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.04 |
| A month of a busy support chatbot | 5M in / 2M out | $0.65 |
Category rankings
Where NVIDIA: Nemotron 3 Nano 30B A3B places across the 5 categories it ranks in. How we rank →
| # | Category | Score |
|---|---|---|
| #14 | Cheap Bulk InferenceCost · of 25 ranked | 137 |
| #15 | Social Media PostsWriting · of 25 ranked | 119 |
| #15 | Voice Assistant BackendVoice · of 25 ranked | 123 |
| #16 | Self-Hosted / LocalCost · of 25 ranked | 117 |
| #22 | Real-Time ChatLatency · of 25 ranked | 117 |
Similar models
NVIDIA: Nemotron 3 Nano Omni (free)
NVIDIA: Nemotron 3 Super
NVIDIA: Nemotron 3 Ultra
NVIDIA: Nemotron 3 Super (free)
NVIDIA: Nemotron Nano 12B 2 VL (free)
NVIDIA: Nemotron 3 Ultra (free)
Quick answers
- How much does NVIDIA: Nemotron 3 Nano 30B A3B cost?
- $0.05 per million input tokens and $0.20 per million output tokens.
- What is NVIDIA: Nemotron 3 Nano 30B A3B's context window?
- 262,144 tokens, roughly 393 pages of text in a single request.
- Does NVIDIA: Nemotron 3 Nano 30B A3B support tool calling?
- It supports tool calling, structured output, a reasoning mode.
- Can NVIDIA: Nemotron 3 Nano 30B A3B process images?
- No, it is text-only on the input side.