25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / Alibaba (Qwen) / Qwen / Qwen3.8 Omni Flash

Qwen3.8 Omni Flash

Qwen omni model for text, vision, audio, and multimodal agent tasks

Line

Qwen

Weights

API only

Released

2026-09-17

Context

1M

Input

$0.1126 / 1M

Output

$0.38 / 1M

Max output

131K

Coverage

1 signal

Overview

Qwen3.8 Omni Flash is an omni model from Alibaba's Qwen line for text, vision, audio, and multimodal agent tasks. It is the successor to Qwen3.8 Flash, released 2026-08-26. It has a context window of 1,000,000 tokens and a max output of 131,072 tokens. Pricing starts at $0.1126 per million input tokens and $0.38 per million output tokens, with cached input at $0.014 per million. Weights are not open. It accepts audio, image, text, and video, and returns text. It supports reasoning mode and tool calling.

Specs

Released2026-09-17
LineAlibaba (Qwen) · Qwen
WeightsAPI only
Context1M tokens
Max output131K tokens
Inputaudio, image, text, video
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoffnot stated
API idqwen3.8-omni-flash

Pricing

Input$0.1126 / 1M tokens
Cached input$0.014 / 1M tokens
Output$0.38 / 1M tokens

Against Qwen3.8 Flash: input $0.11 to $0.1126, output $0.38 to $0.38, cached input $0.011 to $0.014.

Range across the hosts that serve it: up to $0.15 per 1M input tokens.

Strengths

  • Handles text, vision, audio, and video inputs
  • 1M token context window
  • 131K max output tokens
  • Supports reasoning and tool calling
  • Served by 9 hosts

Best for

  • Reach for it for multimodal agent tasks
  • Reach for it for long-context processing up to 1M tokens
  • Reach for it for audio and video understanding
  • Reach for it for reasoning with tool calling

How to access

ProviderModel id
AIHubMixqwen3.8-omni-flash
Alibabaqwen3.8-omni-flash
Alibaba (China)qwen3.8-omni-flash
CrossModelqwen3.8-omni-flash
Eden AIqwen3.8-omni-flash
Kilo Gatewayqwen3.8-omni-flash
NanoGPTqwen3.8-omni-flash
OpenRouterqwen3.8-omni-flash
Vercel AI Gatewayqwen3.8-omni-flash

9 hosts serve this model at the price above · the maker's documentation

What the digest said

1 signal of the archive mention this model, each scored against the published bar. This list is rebuilt from the archive at every build, so it cannot go stale.

Qwen: every version

VersionReleasedContextInput
Qwen3.8 Omni FlashCurrent2026-09-171M$0.1126
Qwen3.8 Flash2026-08-261M$0.11
Qwen3.8 Max2026-08-031M$0.338
Qwen3.7 Flash2026-07-151M$0.0282
Qwen3.7 Plus2026-06-021M$0.282
Qwen3.7 Max2026-05-211M$0.825

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window?
The context window is 1,000,000 tokens.
Are the weights open?
No, the weights are not open.
What input types does it accept?
It accepts audio, image, text, and video.
ProprietaryReasoning1M context