25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / NVIDIA / Nemotron / nvidia-nemotron-nano-9b-v2

nvidia-nemotron-nano-9b-v2

Compact Nemotron model for efficient reasoning and deployable AI agents

Line

Nemotron

Weights

Open

Released

2025-08-18

Context

131K

Input

$0.00 / 1M

Output

$0.00 / 1M

Max output

131K

Coverage

not covered yet

Overview

nvidia-nemotron-nano-9b-v2 is a compact Nemotron model from NVIDIA for efficient reasoning and deployable AI agents. It is the first version of its line in the registry. It has a context window of 131072 tokens and a maximum output of 131072 tokens. The weights are open. It supports reasoning mode and tool calling. It accepts text and returns text. The price is from $0.00 per million input tokens and $0.00 per million output tokens. Knowledge cutoff is 2024-09.

Specs

Released2025-08-18
LineNVIDIA · Nemotron
WeightsOpen
Context131K tokens
Max output131K tokens
Inputtext
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoff2024-09
API idnvidia/nvidia-nemotron-nano-9b-v2

Pricing

Input$0.00 / 1M tokens
Cached inputnot stated
Output$0.00 / 1M tokens

First version of its line in the registry.

Strengths

  • 131072 token context window
  • 131072 token max output
  • Open weights for customization
  • Supports tool calling
  • Reasoning mode enabled

Best for

  • Reach for it for building deployable AI agents
  • Reach for it for long-context text reasoning
  • Reach for it for tool-calling workflows

How to access

ProviderModel id
Nvidianvidia/nvidia-nemotron-nano-9b-v2

1 host serve this model at the price above · the maker's documentation

Nemotron: every version

VersionReleasedContextInput
Nemotron 3 Nano Omni2026-04-28256K$0.10
Nemotron 3 Super2026-03-11262K$0.05
Nemotron Nano 12B v2 VL2025-10-28128K$0.20
nvidia-nemotron-nano-9b-v22025-08-18131K$0.00

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window?
The context window is 131072 tokens, and the maximum output is also 131072 tokens.
Is the model free to use?
The price is from $0.00 per million input tokens and $0.00 per million output tokens.
Does it support tool calling?
Yes, it supports tool calling and reasoning mode.
Open weightsReasoning131K context