25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / Alibaba (Qwen) / Qwen / Qwen3.8 Max

Qwen3.8 Max

2.4-trillion-parameter MoE flagship for coding, professional work, multimodal understanding, and long-horizon agentic workflows

Line

Qwen

Weights

API only

Released

2026-08-03

Context

1M

Input

$0.338 / 1M

Output

$1.01 / 1M

Max output

131K

Coverage

2 signals

Overview

Qwen3.8 Max is a 2.4-trillion-parameter mixture-of-experts flagship from Alibaba's Qwen line, aimed at coding, professional work, multimodal understanding, and long-horizon agentic workflows. It is the latest in the Qwen line, following Qwen3.7 Flash released 2026-07-15. It has a 1,000,000-token context window and a 131,072-token output limit. Pricing starts at $0.338 per million input tokens and $1.01 per million output tokens, with cached input at $0.0676 per million tokens. Weights are not open. It supports reasoning mode and tool calling, accepts image, pdf, text, and video, and returns text.

Specs

Released2026-08-03
LineAlibaba (Qwen) · Qwen
WeightsAPI only
Context1M tokens
Max output131K tokens
Inputimage, pdf, text, video
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoffnot stated
API idqwen3.8-max

Pricing

Input$0.338 / 1M tokens
Cached input$0.0676 / 1M tokens
Output$1.01 / 1M tokens

Against Qwen3.7 Flash: input $0.0282 to $0.338, output $0.1128 to $1.01, cached input $0.003 to $0.0676.

Range across the hosts that serve it: up to $2.50 per 1M input tokens.

Strengths

  • 2.4T parameter MoE for complex tasks
  • 1M token context window
  • 131K token output limit
  • Multimodal input: image, pdf, video
  • Tool calling and reasoning mode

Best for

  • Reach for it for long-horizon agentic workflows
  • Reach for it for coding and professional work
  • Reach for it for multimodal understanding tasks

How to access

ProviderModel id
302.AIqwen3.8-max
AIHubMixqwen3.8-max
Abacusqwen3.8-max
Alibabaqwen3.8-max
Alibaba (China)qwen3.8-max
Alibaba Token Planqwen3.8-max
Alibaba Token Plan (China)qwen3.8-max
Charm Hyperqwen3.8-max
ClinePassqwen3.8-max
Cloudflare AI Gatewayqwen3.8-max
CrossModelqwen3.8-max
Deep Infraqwen3.8-max

31 hosts serve this model at the price above · the maker's documentation

What the digest said

2 signals of the archive mention this model, each scored against the published bar. This list is rebuilt from the archive at every build, so it cannot go stale.

Qwen: every version

VersionReleasedContextInput
Qwen3.8 Omni FlashCurrent2026-09-171M$0.1126
Qwen3.8 Flash2026-08-261M$0.11
Qwen3.8 Max2026-08-031M$0.338
Qwen3.7 Flash2026-07-151M$0.0282
Qwen3.7 Plus2026-06-021M$0.282
Qwen3.7 Max2026-05-211M$0.825

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window size?
1,000,000 tokens.
Is it open weights?
No, weights are not open.
What input types does it accept?
Image, pdf, text, and video.
ProprietaryReasoning1M context