Model registry / Alibaba (Qwen) / Qwen / Qwen3.8 Omni Flash
Qwen3.8 Omni Flash
Qwen omni model for text, vision, audio, and multimodal agent tasks
Line
Qwen
Weights
API only
Released
2026-09-17
Context
1M
Input
$0.1126 / 1M
Output
$0.38 / 1M
Max output
131K
Coverage
1 signal
Overview
Qwen3.8 Omni Flash is an omni model from Alibaba's Qwen line for text, vision, audio, and multimodal agent tasks. It is the successor to Qwen3.8 Flash, released 2026-08-26. It has a context window of 1,000,000 tokens and a max output of 131,072 tokens. Pricing starts at $0.1126 per million input tokens and $0.38 per million output tokens, with cached input at $0.014 per million. Weights are not open. It accepts audio, image, text, and video, and returns text. It supports reasoning mode and tool calling.
Specs
Pricing
Against Qwen3.8 Flash: input $0.11 to $0.1126, output $0.38 to $0.38, cached input $0.011 to $0.014.
Range across the hosts that serve it: up to $0.15 per 1M input tokens.
Strengths
- Handles text, vision, audio, and video inputs
- 1M token context window
- 131K max output tokens
- Supports reasoning and tool calling
- Served by 9 hosts
Best for
- Reach for it for multimodal agent tasks
- Reach for it for long-context processing up to 1M tokens
- Reach for it for audio and video understanding
- Reach for it for reasoning with tool calling
How to access
9 hosts serve this model at the price above · the maker's documentation
What the digest said
1 signal of the archive mention this model, each scored against the published bar. This list is rebuilt from the archive at every build, so it cannot go stale.
Qwen: every version
The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.
FAQ
- What is the context window?
- The context window is 1,000,000 tokens.
- Are the weights open?
- No, the weights are not open.
- What input types does it accept?
- It accepts audio, image, text, and video.