Model registry / Z.ai (Zhipu, GLM) / GLM / GLM-5.3
GLM-5.3
Flagship GLM model for long-horizon coding, agents, and complex project delivery
Line
GLM
Weights
Open
Released
2026-08-14
Context
1M
Input
$0.40 / 1M
Output
$1.40 / 1M
Max output
131K
Coverage
not covered yet
Overview
GLM-5.3 is the flagship model in the GLM line from Z.ai (Zhipu). It is designed for long-horizon coding, agents, and complex project delivery. It has a context window of 1,000,000 tokens and a maximum output of 131,072 tokens. It accepts and returns text. It supports reasoning mode and tool calling. The weights are open. Pricing starts at $0.40 per million input tokens and $1.40 per million output tokens, with cached input at $0.06 per million tokens.
Specs
Pricing
Against GLM-5.2: input $0.30 to $0.40, output $1.05 to $1.40, cached input $0.05 to $0.06.
Range across the hosts that serve it: up to $2.80 per 1M input tokens.
Strengths
- 1,000,000 token context window
- 131,072 token max output
- Supports reasoning mode and tool calling
- Open weights for self-hosting
- Built for long-horizon coding and agents
Best for
- Reach for it for long-horizon coding projects
- Reach for it for building autonomous agents
- Reach for it for complex project delivery
- Reach for it for tasks needing a million-token context
How to access
57 hosts serve this model at the price above · the maker's documentation
GLM: every version
The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.
FAQ
- How large is the context window?
- The context window is 1,000,000 tokens, and the maximum output is 131,072 tokens.
- Is the model open weights?
- Yes, the weights are open.
- What is the pricing?
- Input from $0.40 per million tokens, output from $1.40 per million tokens, cached input from $0.06 per million tokens.