Model registry / Alibaba (Qwen) / Qwen / Qwen3.7 Flash
Qwen3.7 Flash
Lightweight multimodal Qwen model for high-throughput text, image, and video tasks
Line
Qwen
Weights
API only
Released
2026-07-15
Context
1M
Input
$0.0282 / 1M
Output
$0.1128 / 1M
Max output
131K
Coverage
not covered yet
Overview
Qwen3.7 Flash is a lightweight multimodal Qwen model for high-throughput text, image, and video tasks. It is part of the Qwen line, released on 2026-07-15, following Qwen3.7 Plus. It has a context window of 1,000,000 tokens and a maximum output of 131,072 tokens. It accepts image, text, and video inputs and returns text. It supports reasoning mode and tool calling. The weights are not open. Price starts at $0.0282 per million input tokens and $0.1128 per million output tokens, with cached input from $0.003 per million tokens.
Specs
Pricing
Against Qwen3.7 Plus: input $0.282 to $0.0282, output $1.13 to $0.1128, cached input $0.032 to $0.003.
Range across the hosts that serve it: up to $0.20 per 1M input tokens.
Strengths
- 1M token context window
- 131K token max output
- Accepts image, text, video
- High-throughput design
- Supports reasoning and tool calling
Best for
- Reach for it for high-throughput text, image, and video tasks
- Reach for it for long-context processing up to 1M tokens
- Reach for it for multimodal input with text output
- Reach for it for cost-sensitive deployments with low input price
How to access
12 hosts serve this model at the price above · the maker's documentation
Qwen: every version
The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.
FAQ
- What is the context window?
- The context window is 1,000,000 tokens.
- Does it support tool calling?
- Yes, it supports tool calling and reasoning mode.
- Are the weights open?
- No, the weights are not open.