Model registry / Z.ai (Zhipu, GLM) / GLM / GLM-5
GLM-5
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
Line
GLM
Weights
Open
Released
2026-02-12
Context
205K
Input
$0.50 / 1M
Output
$1.92 / 1M
Max output
131K
Coverage
not covered yet
Overview
GLM-5 is a general GLM flagship for coding, analysis, and tool-heavy engineering workflows. It is the first version of its line in the registry, released on 2026-02-12. It has a context window of 204800 tokens and a maximum output of 131072 tokens. It accepts text and returns text, supports reasoning mode and tool calling, and has open weights. Price starts at $0.50 per million input tokens and $1.92 per million output tokens, with cached input from $0.10 per million tokens.
Specs
Pricing
First version of its line in the registry.
Range across the hosts that serve it: up to $1.10 per 1M input tokens.
Strengths
- 204800 token context window
- 131072 token max output
- Open weights
- Reasoning mode
- Tool calling
Best for
- Reach for it for coding
- Reach for it for analysis
- Reach for it for tool-heavy engineering workflows
How to access
42 hosts serve this model at the price above · the maker's documentation
GLM: every version
The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.
FAQ
- What is the context window size?
- The context window is 204800 tokens, and the maximum output is 131072 tokens.
- Are the weights open?
- Yes, the weights are open. The model also supports reasoning mode and tool calling.
- What is the pricing?
- Input tokens cost from $0.50 per million, output from $1.92 per million, and cached input from $0.10 per million.