Model registry / Z.ai (Zhipu, GLM) / GLM / GLM-5V-Turbo
GLM-5V-Turbo
Fast GLM vision model for screenshots, documents, and multimodal agent tasks
Line
GLM
Weights
API only
Released
2026-04-01
Context
200K
Input
$0.7042 / 1M
Output
$3.10 / 1M
Max output
131K
Coverage
not covered yet
Overview
GLM-5V-Turbo is a vision model from Z.ai (Zhipu, GLM) for screenshots, documents, and multimodal agent tasks. It is part of the GLM line, released on 2026-04-01, following GLM-5-Turbo. It has a context window of 200,000 tokens and a maximum output of 131,072 tokens. It accepts image, PDF, text, and video inputs, and returns text. It supports reasoning mode and tool calling. Weights are not open. Pricing starts at $0.7042 per million input tokens and $3.10 per million output tokens, with cached input at $0.169 per million tokens. It is served by 17 hosts.
Specs
Pricing
Against GLM-5-Turbo: input $0.72 to $0.7042, output $3.19 to $3.10, cached input $0.174 to $0.169.
Range across the hosts that serve it: up to $5.00 per 1M input tokens.
Strengths
- Fast vision model for screenshots and documents
- Accepts image, PDF, text, and video inputs
- 200K token context window
- 131K token max output
- Supports reasoning mode and tool calling
Best for
- Reach for it for screenshot analysis
- Reach for it for document understanding
- Reach for it for multimodal agent tasks
How to access
17 hosts serve this model at the price above · the maker's documentation
GLM: every version
The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.
FAQ
- What input types does GLM-5V-Turbo support?
- It accepts image, PDF, text, and video, and returns text.
- Is GLM-5V-Turbo open source?
- No, the weights are not open.
- What is the pricing for GLM-5V-Turbo?
- Input from $0.7042 per million tokens, output $3.10 per million, cached input $0.169 per million.