25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / NVIDIA / Nemotron / Nemotron 3.5 Lightning 30B A3B

Nemotron 3.5 Lightning 30B A3B

Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads

Line

Nemotron

Weights

Open

Released

2026-08-11

Context

262K

Input

$0.00 / 1M

Output

$0.00 / 1M

Max output

262K

Coverage

not covered yet

Overview

Nemotron 3.5 Lightning 30B A3B is a fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads. It is part of the Nemotron line, succeeding the Nemotron 3 Ultra 550B A55B released 2026-06-04. It has a context window of 262144 tokens and a max output of 262144 tokens. The weights are open. It supports reasoning mode and tool calling. It accepts text and returns text. The price is from $0.00 per million input tokens and $0.00 per million output tokens.

Specs

Released2026-08-11
LineNVIDIA · Nemotron
WeightsOpen
Context262K tokens
Max output262K tokens
Inputtext
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoffnot stated
API idnvidia/nemotron-3.5-lightning-30b-a3b

Pricing

Input$0.00 / 1M tokens
Cached inputnot stated
Output$0.00 / 1M tokens

Against Nemotron 3 Ultra 550B A55B: window 1M to 262K, input $0.50 to $0.00, output $2.20 to $0.00.

Strengths

  • Fast MoE architecture for agentic tasks
  • 262k token context window
  • 262k token max output
  • Open weights
  • Supports reasoning and tool calling

Best for

  • Reach for it for reliable agentic workflows
  • Reach for it for enterprise workload automation
  • Reach for it for long-context reasoning tasks
  • Reach for it for tool-calling agents

How to access

ProviderModel id
Merge Gatewaynvidia/nemotron-3.5-lightning-30b-a3b
Nvidianvidia/nemotron-3.5-lightning-30b-a3b
Requestynvidia/nemotron-3.5-lightning-30b-a3b

3 hosts serve this model at the price above · the maker's documentation

Nemotron: every version

VersionReleasedContextInput
Nemotron 3 Nano Omni2026-04-28256K$0.10
Nemotron 3 Super2026-03-11262K$0.05
Nemotron Nano 12B v2 VL2025-10-28128K$0.20
nvidia-nemotron-nano-9b-v22025-08-18131K$0.00

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window and max output?
The context window is 262144 tokens, and the max output is also 262144 tokens.
Are the weights open?
Yes, the weights are open.
What is the price?
The price is from $0.00 per million input tokens and $0.00 per million output tokens.
Open weightsReasoning262K context