25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / Alibaba (Qwen) / Qwen / Qwen3.7 Flash

Qwen3.7 Flash

Lightweight multimodal Qwen model for high-throughput text, image, and video tasks

Line

Qwen

Weights

API only

Released

2026-07-15

Context

1M

Input

$0.0282 / 1M

Output

$0.1128 / 1M

Max output

131K

Coverage

not covered yet

Overview

Qwen3.7 Flash is a lightweight multimodal Qwen model for high-throughput text, image, and video tasks. It is part of the Qwen line, released on 2026-07-15, following Qwen3.7 Plus. It has a context window of 1,000,000 tokens and a maximum output of 131,072 tokens. It accepts image, text, and video inputs and returns text. It supports reasoning mode and tool calling. The weights are not open. Price starts at $0.0282 per million input tokens and $0.1128 per million output tokens, with cached input from $0.003 per million tokens.

Specs

Released2026-07-15
LineAlibaba (Qwen) · Qwen
WeightsAPI only
Context1M tokens
Max output131K tokens
Inputimage, text, video
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoffnot stated
API idqwen3.7-flash

Pricing

Input$0.0282 / 1M tokens
Cached input$0.003 / 1M tokens
Output$0.1128 / 1M tokens

Against Qwen3.7 Plus: input $0.282 to $0.0282, output $1.13 to $0.1128, cached input $0.032 to $0.003.

Range across the hosts that serve it: up to $0.20 per 1M input tokens.

Strengths

  • 1M token context window
  • 131K token max output
  • Accepts image, text, video
  • High-throughput design
  • Supports reasoning and tool calling

Best for

  • Reach for it for high-throughput text, image, and video tasks
  • Reach for it for long-context processing up to 1M tokens
  • Reach for it for multimodal input with text output
  • Reach for it for cost-sensitive deployments with low input price

How to access

ProviderModel id
AIHubMixqwen3.7-flash
Alibabaqwen3.7-flash
Alibaba (China)qwen3.7-flash
Charm Hyperqwen3.7-flash
CrossModelqwen3.7-flash
DevPass (LLM Gateway)qwen3.7-flash
Kilo Gatewayqwen3.7-flash
LLM Gatewayqwen3.7-flash
NanoGPTqwen3.7-flash
OpenRouterqwen3.7-flash
OrcaRouterqwen3.7-flash
Vercel AI Gatewayqwen3.7-flash

12 hosts serve this model at the price above · the maker's documentation

Qwen: every version

VersionReleasedContextInput
Qwen3.8 Omni FlashCurrent2026-09-171M$0.1126
Qwen3.8 Flash2026-08-261M$0.11
Qwen3.8 Max2026-08-031M$0.338
Qwen3.7 Flash2026-07-151M$0.0282
Qwen3.7 Plus2026-06-021M$0.282
Qwen3.7 Max2026-05-211M$0.825

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window?
The context window is 1,000,000 tokens.
Does it support tool calling?
Yes, it supports tool calling and reasoning mode.
Are the weights open?
No, the weights are not open.
ProprietaryReasoning1M context