Model registry / Alibaba (Qwen) / Qwen3 6 / Qwen3.6 Flash
Qwen3.6 Flash
Qwen vision-language model for visual reasoning, documents, and agent tasks
Line
Qwen3 6
Weights
API only
Released
2026-04-27
Context
1M
Input
$0.165 / 1M
Output
$0.99 / 1M
Max output
66K
Coverage
not covered yet
Overview
Qwen3.6 Flash is a vision-language model from Alibaba's Qwen line for visual reasoning, documents, and agent tasks. It is the first version of the qwen3.6 line in the registry. It has a context window of 1,000,000 tokens and a maximum output of 65,536 tokens. Weights are not open. It supports reasoning mode and tool calling. It accepts image, text, and video input and returns text. Pricing starts at $0.165 per million input tokens and $0.99 per million output tokens, with cached input at $0.0169 per million tokens.
Specs
Pricing
First version of its line in the registry.
Range across the hosts that serve it: up to $0.25 per 1M input tokens.
Strengths
- 1,000,000 token context window
- 65,536 token maximum output
- Accepts image, text, and video input
- Supports reasoning mode and tool calling
Best for
- Reach for it for visual reasoning tasks
- Reach for it for document analysis
- Reach for it for agent tasks with tool calling
How to access
18 hosts serve this model at the price above · the maker's documentation
Qwen3 6: every version
The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.
FAQ
- What is the context window size?
- The context window is 1,000,000 tokens, and the maximum output is 65,536 tokens.
- Does it support tool calling?
- Yes, it supports tool calling and also has a reasoning mode.
- What input types does it accept?
- It accepts image, text, and video input, and returns text output.