nvidia

NVIDIA: Nemotron 3 Nano 30B A3B

The NVIDIA Nemotron 3 Nano 30B A3B is designed for handling extensive text inputs with a context length of up to 262,144 tokens and can generate up to 228,000 output tokens. It supports tools and reasoning, making it suitable for complex tasks that require structured data processing. This model stands out with its comprehensive input modality focusing on text alone. For those prioritizing a blend of performance and cost-effectiveness, the Nemotron 3 Nano 30B A3B is well-worth considering. With a blended benchmark score of 10.7 from a single independent source and pricing at $0.05 per million input tokens and $0.2 per million output tokens, it offers competitive value for its capabilities. While the benchmark coverage is comprehensive, the model's specific use case should be evaluated against your needs to ensure optimal performance and cost efficiency.

Quality Score
99/100
price + capability + benchmarks
Input Price
$0.05
per 1M tokens
Output Price
$0.20
per 1M tokens
Context Window
262,144
tokens

Benchmark results

Independent, published benchmarks. Blended score 10.6 across 1 benchmark, last refreshed 2026-07-27. How scoring works →

BenchmarkMeasuresScore
AI Index broad capability composite 12.2
Model ID
nvidia/nemotron-3-nano-30b-a3b
Vendor
nvidia
Released
December 2025
Tokenizer
Other
Input Modalities
text
Output Modalities
text
Max Output
228,000 tokens
Tool Calling
✓ supported
Structured Output
✓ supported
Reasoning Mode
✓ supported
Vision
text only
Audio
no
Moderated
no

What it costs in practice

Computed from the current $0.05/M input and $0.20/M output rates. Run your own numbers →

JobTokensCost
Summarize a 50-page report 30k in / 1.5k out under $0.01
Classify 1,000 customer emails 500k in / 50k out $0.04
A month of a busy support chatbot 5M in / 2M out $0.65

Category rankings

Where NVIDIA: Nemotron 3 Nano 30B A3B places across the 5 categories it ranks in. How we rank →

#CategoryScore
#14 Cheap Bulk InferenceCost · of 25 ranked 137
#15 Social Media PostsWriting · of 25 ranked 119
#15 Voice Assistant BackendVoice · of 25 ranked 123
#16 Self-Hosted / LocalCost · of 25 ranked 117
#22 Real-Time ChatLatency · of 25 ranked 117

Similar models

Quick answers

How much does NVIDIA: Nemotron 3 Nano 30B A3B cost?
$0.05 per million input tokens and $0.20 per million output tokens.
What is NVIDIA: Nemotron 3 Nano 30B A3B's context window?
262,144 tokens, roughly 393 pages of text in a single request.
Does NVIDIA: Nemotron 3 Nano 30B A3B support tool calling?
It supports tool calling, structured output, a reasoning mode.
Can NVIDIA: Nemotron 3 Nano 30B A3B process images?
No, it is text-only on the input side.