25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / Alibaba (Qwen) / Qwen / Qwen3.8 Flash

Qwen3.8 Flash

Qwen vision-language model for visual reasoning, documents, and agent tasks

Line

Qwen

Weights

API only

Released

2026-08-26

Context

1M

Input

$0.11 / 1M

Output

$0.38 / 1M

Max output

131K

Coverage

not covered yet

Overview

Qwen3.8 Flash is a vision-language model from Alibaba's Qwen line, designed for visual reasoning, documents, and agent tasks. It is a later release in the Qwen line, following Qwen3.8 Max. It has a context window of 1,000,000 tokens and a maximum output of 131,072 tokens. Pricing starts at $0.11 per million input tokens and $0.38 per million output tokens, with cached input at $0.011 per million tokens. The weights are not open. It supports reasoning mode, tool calling, and accepts image, text, and video input, returning text output.

Specs

Released2026-08-26
LineAlibaba (Qwen) · Qwen
WeightsAPI only
Context1M tokens
Max output131K tokens
Inputimage, text, video
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoffnot stated
API idqwen3.8-flash

Pricing

Input$0.11 / 1M tokens
Cached input$0.011 / 1M tokens
Output$0.38 / 1M tokens

Against Qwen3.8 Max: input $0.338 to $0.11, output $1.01 to $0.38, cached input $0.0676 to $0.011.

Range across the hosts that serve it: up to $0.18 per 1M input tokens.

Strengths

  • 1M token context window
  • 131K max output tokens
  • Accepts image, text, and video
  • Reasoning mode and tool calling
  • Priced from $0.11/M input tokens

Best for

  • Reach for it for visual reasoning tasks
  • Reach for it for document analysis
  • Reach for it for agent tasks with tool calling

How to access

ProviderModel id
302.AIqwen3.8-flash
AIHubMixqwen3.8-flash
Alibabaqwen3.8-flash
Alibaba (China)qwen3.8-flash
Alibaba Token Planqwen3.8-flash
Alibaba Token Plan (China)qwen3.8-flash
Charm Hyperqwen3.8-flash
CrossModelqwen3.8-flash
Deep Infraqwen3.8-flash
DevPass (LLM Gateway)qwen3.8-flash
Eden AIqwen3.8-flash
GMI Cloudqwen3.8-flash

25 hosts serve this model at the price above · the maker's documentation

Qwen: every version

VersionReleasedContextInput
Qwen3.8 Omni FlashCurrent2026-09-171M$0.1126
Qwen3.8 Flash2026-08-261M$0.11
Qwen3.8 Max2026-08-031M$0.338
Qwen3.7 Flash2026-07-151M$0.0282
Qwen3.7 Plus2026-06-021M$0.282
Qwen3.7 Max2026-05-211M$0.825

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window?
The context window is 1,000,000 tokens.
Does it support tool calling?
Yes, it supports tool calling and reasoning mode.
Is it open weights?
No, the weights are not open.
ProprietaryReasoning1M context