Model registry / NVIDIA / Nemotron / Nemotron 3.5 Lightning 30B A3B
Nemotron 3.5 Lightning 30B A3B
Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
Line
Nemotron
Weights
Open
Released
2026-08-11
Context
262K
Input
$0.00 / 1M
Output
$0.00 / 1M
Max output
262K
Coverage
not covered yet
Overview
Nemotron 3.5 Lightning 30B A3B is a fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads. It is part of the Nemotron line, succeeding the Nemotron 3 Ultra 550B A55B released 2026-06-04. It has a context window of 262144 tokens and a max output of 262144 tokens. The weights are open. It supports reasoning mode and tool calling. It accepts text and returns text. The price is from $0.00 per million input tokens and $0.00 per million output tokens.
Specs
Pricing
Against Nemotron 3 Ultra 550B A55B: window 1M to 262K, input $0.50 to $0.00, output $2.20 to $0.00.
Strengths
- Fast MoE architecture for agentic tasks
- 262k token context window
- 262k token max output
- Open weights
- Supports reasoning and tool calling
Best for
- Reach for it for reliable agentic workflows
- Reach for it for enterprise workload automation
- Reach for it for long-context reasoning tasks
- Reach for it for tool-calling agents
How to access
3 hosts serve this model at the price above · the maker's documentation
Nemotron: every version
The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.
FAQ
- What is the context window and max output?
- The context window is 262144 tokens, and the max output is also 262144 tokens.
- Are the weights open?
- Yes, the weights are open.
- What is the price?
- The price is from $0.00 per million input tokens and $0.00 per million output tokens.