nvidia

NVIDIA: Nemotron 3 Nano 30B A3B (free)

NVIDIA's Nemotron 3 Nano 30B A3B is designed for handling extensive text inputs with a context length of up to 256,000 tokens, making it suitable for large-scale applications and complex queries. The model supports input modalities solely in text and offers reasoning capabilities alongside the ability to interface with external tools. While structured output support is not specified, its robust reasoning skills ensure effective problem-solving. Considering its free price point and benchmark score of 10.4 across a single independent test, this model stands as an excellent choice for those prioritizing cost efficiency without compromising on performance. Its suitability spans from educational projects to small-scale enterprise applications where budget constraints are significant but high-quality results remain essential.

Quality Score
88/100
price + capability + benchmarks
Input Price
Free
per 1M tokens
Output Price
Free
per 1M tokens
Context Window
256,000
tokens

Benchmark results

Independent, published benchmarks. Blended score 10.4 across 1 benchmark, last refreshed 2026-08-13. How scoring works →

BenchmarkMeasuresScore
AI Index broad capability composite 11.8
Model ID
nvidia/nemotron-3-nano-30b-a3b:free
Vendor
nvidia
Released
December 2025
Tokenizer
Other
Input Modalities
text
Output Modalities
text
Max Output
default
Tool Calling
✓ supported
Structured Output
not supported
Reasoning Mode
✓ supported
Vision
text only
Audio
no
Moderated
no

Similar models

Quick answers

Is NVIDIA: Nemotron 3 Nano 30B A3B (free) free to use?
Yes. It is currently listed at no cost per token via OpenRouter, subject to provider rate limits.
What is NVIDIA: Nemotron 3 Nano 30B A3B (free)'s context window?
256,000 tokens, roughly 384 pages of text in a single request.
Does NVIDIA: Nemotron 3 Nano 30B A3B (free) support tool calling?
It supports tool calling, a reasoning mode.
Can NVIDIA: Nemotron 3 Nano 30B A3B (free) process images?
No, it is text-only on the input side.