Model registry / DeepSeek / Deepseek Thinking / DeepSeek V4 Pro
DeepSeek V4 Pro
DeepSeek V4 Pro snapshot with million-token context and support for thinking and non-thinking modes
Line
Deepseek Thinking
Weights
Open
Released
2026-08-12
Context
1M
Input
$0.2088 / 1M
Output
$0.4176 / 1M
Max output
393K
Coverage
not covered yet
Overview
DeepSeek V4 Pro is a snapshot in the deepseek-thinking line, released on 2026-08-12. It supports thinking and non-thinking modes and offers a million-token context window. This is the first version of its line in the registry. The model has a context window of 1,000,000 tokens and a maximum output of 393,216 tokens. It accepts text and returns text. It supports reasoning mode and tool calling. The weights are open. Pricing starts at $0.2088 per million input tokens and $0.4176 per million output tokens, with cached input from $0.003 per million tokens. It is served by 66 hosts.
Specs
Pricing
First version of its line in the registry.
Range across the hosts that serve it: up to $2.40 per 1M input tokens.
Strengths
- Million-token context window
- Supports thinking and non-thinking modes
- Open weights
- Tool calling support
- 393,216 token max output
Best for
- Reach for it for long-document analysis
- Reach for it for complex reasoning with thinking mode
- Reach for it for tool-calling workflows
- Reach for it for processing large contexts
How to access
66 hosts serve this model at the price above · the maker's documentation
Deepseek Thinking: every version
The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.
FAQ
- What is the context window?
- The context window is 1,000,000 tokens, and the maximum output is 393,216 tokens.
- Are the weights open?
- Yes, the weights are open. The model also supports reasoning mode and tool calling.
- What does it cost?
- Input starts at $0.2088 per million tokens, output at $0.4176 per million tokens. Cached input is from $0.003 per million tokens.