Model registry / NVIDIA / Nemotron / nvidia-nemotron-nano-9b-v2
nvidia-nemotron-nano-9b-v2
Compact Nemotron model for efficient reasoning and deployable AI agents
Line
Nemotron
Weights
Open
Released
2025-08-18
Context
131K
Input
$0.00 / 1M
Output
$0.00 / 1M
Max output
131K
Coverage
not covered yet
Overview
nvidia-nemotron-nano-9b-v2 is a compact Nemotron model from NVIDIA for efficient reasoning and deployable AI agents. It is the first version of its line in the registry. It has a context window of 131072 tokens and a maximum output of 131072 tokens. The weights are open. It supports reasoning mode and tool calling. It accepts text and returns text. The price is from $0.00 per million input tokens and $0.00 per million output tokens. Knowledge cutoff is 2024-09.
Specs
Pricing
First version of its line in the registry.
Strengths
- 131072 token context window
- 131072 token max output
- Open weights for customization
- Supports tool calling
- Reasoning mode enabled
Best for
- Reach for it for building deployable AI agents
- Reach for it for long-context text reasoning
- Reach for it for tool-calling workflows
How to access
1 host serve this model at the price above · the maker's documentation
Nemotron: every version
The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.
FAQ
- What is the context window?
- The context window is 131072 tokens, and the maximum output is also 131072 tokens.
- Is the model free to use?
- The price is from $0.00 per million input tokens and $0.00 per million output tokens.
- Does it support tool calling?
- Yes, it supports tool calling and reasoning mode.