25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / NVIDIA / Nemotron / Nemotron 3 Nano Omni

Nemotron 3 Nano Omni

Open Nemotron omni model combining reasoning with text, vision, and audio

Line

Nemotron

Weights

Open

Released

2026-04-28

Context

256K

Input

$0.10 / 1M

Output

$0.25 / 1M

Max output

66K

Coverage

not covered yet

Overview

Nemotron 3 Nano Omni is an open omni model from NVIDIA that combines reasoning with text, vision, and audio. It is part of the Nemotron line and was released on 2026-04-28. It has a context window of 256,000 tokens and a maximum output of 65,536 tokens. The weights are open. It supports reasoning mode and tool calling. It accepts audio, image, text, and video, and returns text. Pricing starts at $0.10 per million input tokens and $0.25 per million output tokens. It is served by 6 hosts.

Specs

Released2026-04-28
LineNVIDIA · Nemotron
WeightsOpen
Context256K tokens
Max output66K tokens
Inputaudio, image, text, video
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoffnot stated
API idnvidia/nemotron-3-nano-omni-30b-a3b-reasoning

Pricing

Input$0.10 / 1M tokens
Cached inputnot stated
Output$0.25 / 1M tokens

Against Nemotron 3 Super: window 262K to 256K, input $0.05 to $0.10.

Range across the hosts that serve it: up to $0.20 per 1M input tokens.

Strengths

  • 256K token context window
  • 65K token max output
  • Open weights
  • Reasoning mode and tool calling
  • Accepts audio, image, text, video

Best for

  • Reach for it for multimodal reasoning with text, vision, audio, and video
  • Reach for it for long-context tasks up to 256K tokens
  • Reach for it for tool-calling agents with open weights

How to access

ProviderModel id
Deep Infranvidia/nemotron-3-nano-omni-30b-a3b-reasoning
Kilo Gatewaynvidia/nemotron-3-nano-omni-30b-a3b-reasoning
Nvidianvidia/nemotron-3-nano-omni-30b-a3b-reasoning
OpenRouternvidia/nemotron-3-nano-omni-30b-a3b-reasoning
Requestynvidia/nemotron-3-nano-omni-30b-a3b-reasoning
Vultrnvidia/nemotron-3-nano-omni-30b-a3b-reasoning

6 hosts serve this model at the price above · the maker's documentation

Nemotron: every version

VersionReleasedContextInput
Nemotron 3 Nano Omni2026-04-28256K$0.10
Nemotron 3 Super2026-03-11262K$0.05
Nemotron Nano 12B v2 VL2025-10-28128K$0.20
nvidia-nemotron-nano-9b-v22025-08-18131K$0.00

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window?
256,000 tokens.
What is the price?
From $0.10 per million input tokens and $0.25 per million output tokens.
What modalities does it accept?
Audio, image, text, and video.
Open weightsReasoning256K context