nvidia

NVIDIA: Nemotron 3 Nano 30B A3B

The NVIDIA Nemotron 3 Nano 30B A3B is designed for handling extensive text inputs with a context length of 262,144 tokens and can generate up to 228,000 completion tokens. It supports both reasoning tasks and integration with tools, making it versatile for complex problem-solving scenarios. However, this model lacks structured output capabilities, limiting its suitability for data-driven applications. Considering its price at $0.05 per million input tokens and $0.2 per million output tokens, the Nemotron 3 Nano 30B A3B is more suitable for businesses or individuals with specific text-based needs who can afford the cost but do not require structured outputs. Its benchmark score of 10.4 across a comprehensive set of benchmarks positions it well in terms of performance and coverage; thus, it should be on your shortlist if you prioritize efficiency and affordability in text handling tasks.

Quality Score
99/100
price + capability + benchmarks
Input Price
$0.05
per 1M tokens
Output Price
$0.20
per 1M tokens
Context Window
262,144
tokens

Benchmark results

Independent, published benchmarks. Blended score 10.4 across 1 benchmark, last refreshed 2026-08-13. How scoring works →

BenchmarkMeasuresScore
AI Index broad capability composite 11.8
Model ID
nvidia/nemotron-3-nano-30b-a3b
Vendor
nvidia
Released
December 2025
Tokenizer
Other
Input Modalities
text
Output Modalities
text
Max Output
228,000 tokens
Tool Calling
✓ supported
Structured Output
✓ supported
Reasoning Mode
✓ supported
Vision
text only
Audio
no
Moderated
no

What it costs in practice

Computed from the current $0.05/M input and $0.20/M output rates. Run your own numbers →

JobTokensCost
Summarize a 50-page report 30k in / 1.5k out under $0.01
Classify 1,000 customer emails 500k in / 50k out $0.04
A month of a busy support chatbot 5M in / 2M out $0.65

Category rankings

Where NVIDIA: Nemotron 3 Nano 30B A3B places across the 3 categories it ranks in. How we rank →

#CategoryScore
#20 Cheap Bulk InferenceCost · of 25 ranked 137
#25 Social Media PostsWriting · of 25 ranked 119
#25 Voice Assistant BackendVoice · of 25 ranked 123

Similar models

Quick answers

How much does NVIDIA: Nemotron 3 Nano 30B A3B cost?
$0.05 per million input tokens and $0.20 per million output tokens.
What is NVIDIA: Nemotron 3 Nano 30B A3B's context window?
262,144 tokens, roughly 393 pages of text in a single request.
Does NVIDIA: Nemotron 3 Nano 30B A3B support tool calling?
It supports tool calling, structured output, a reasoning mode.
Can NVIDIA: Nemotron 3 Nano 30B A3B process images?
No, it is text-only on the input side.