25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / Z.ai (Zhipu, GLM) / GLM / GLM-5V-Turbo

GLM-5V-Turbo

Fast GLM vision model for screenshots, documents, and multimodal agent tasks

Line

GLM

Weights

API only

Released

2026-04-01

Context

200K

Input

$0.7042 / 1M

Output

$3.10 / 1M

Max output

131K

Coverage

not covered yet

Overview

GLM-5V-Turbo is a vision model from Z.ai (Zhipu, GLM) for screenshots, documents, and multimodal agent tasks. It is part of the GLM line, released on 2026-04-01, following GLM-5-Turbo. It has a context window of 200,000 tokens and a maximum output of 131,072 tokens. It accepts image, PDF, text, and video inputs, and returns text. It supports reasoning mode and tool calling. Weights are not open. Pricing starts at $0.7042 per million input tokens and $3.10 per million output tokens, with cached input at $0.169 per million tokens. It is served by 17 hosts.

Specs

Released2026-04-01
LineZ.ai (Zhipu, GLM) · GLM
WeightsAPI only
Context200K tokens
Max output131K tokens
Inputimage, pdf, text, video
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoffnot stated
API idglm-5v-turbo

Pricing

Input$0.7042 / 1M tokens
Cached input$0.169 / 1M tokens
Output$3.10 / 1M tokens

Against GLM-5-Turbo: input $0.72 to $0.7042, output $3.19 to $3.10, cached input $0.174 to $0.169.

Range across the hosts that serve it: up to $5.00 per 1M input tokens.

Strengths

  • Fast vision model for screenshots and documents
  • Accepts image, PDF, text, and video inputs
  • 200K token context window
  • 131K token max output
  • Supports reasoning mode and tool calling

Best for

  • Reach for it for screenshot analysis
  • Reach for it for document understanding
  • Reach for it for multimodal agent tasks

How to access

ProviderModel id
302.AIglm-5v-turbo
AIHubMixglm-5v-turbo
Cortecsglm-5v-turbo
DevPass (LLM Gateway)glm-5v-turbo
Eden AIglm-5v-turbo
Kilo Gatewayglm-5v-turbo
LLM Gatewayglm-5v-turbo
NanoGPTglm-5v-turbo
Ofoxglm-5v-turbo
OpenRouterglm-5v-turbo
SiliconFlowglm-5v-turbo
Tempr Gatewayglm-5v-turbo

17 hosts serve this model at the price above · the maker's documentation

GLM: every version

VersionReleasedContextInput
GLM-5.3Current2026-08-141M$0.40
GLM-5.22026-06-131M$0.30
GLM-5.12026-04-07200K$0.45
GLM-5V-Turbo2026-04-01200K$0.7042
GLM-5-Turbo2026-03-16200K$0.72
GLM-52026-02-12205K$0.50

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What input types does GLM-5V-Turbo support?
It accepts image, PDF, text, and video, and returns text.
Is GLM-5V-Turbo open source?
No, the weights are not open.
What is the pricing for GLM-5V-Turbo?
Input from $0.7042 per million tokens, output $3.10 per million, cached input $0.169 per million.
ProprietaryReasoning200K context