25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / Z.ai (Zhipu, GLM) / GLM / GLM-5.3

GLM-5.3

Flagship GLM model for long-horizon coding, agents, and complex project delivery

Line

GLM

Weights

Open

Released

2026-08-14

Context

1M

Input

$0.40 / 1M

Output

$1.40 / 1M

Max output

131K

Coverage

not covered yet

Overview

GLM-5.3 is the flagship model in the GLM line from Z.ai (Zhipu). It is designed for long-horizon coding, agents, and complex project delivery. It has a context window of 1,000,000 tokens and a maximum output of 131,072 tokens. It accepts and returns text. It supports reasoning mode and tool calling. The weights are open. Pricing starts at $0.40 per million input tokens and $1.40 per million output tokens, with cached input at $0.06 per million tokens.

Specs

Released2026-08-14
LineZ.ai (Zhipu, GLM) · GLM
WeightsOpen
Context1M tokens
Max output131K tokens
Inputtext
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoffnot stated
API idglm-5.3

Pricing

Input$0.40 / 1M tokens
Cached input$0.06 / 1M tokens
Output$1.40 / 1M tokens

Against GLM-5.2: input $0.30 to $0.40, output $1.05 to $1.40, cached input $0.05 to $0.06.

Range across the hosts that serve it: up to $2.80 per 1M input tokens.

Strengths

  • 1,000,000 token context window
  • 131,072 token max output
  • Supports reasoning mode and tool calling
  • Open weights for self-hosting
  • Built for long-horizon coding and agents

Best for

  • Reach for it for long-horizon coding projects
  • Reach for it for building autonomous agents
  • Reach for it for complex project delivery
  • Reach for it for tasks needing a million-token context

How to access

ProviderModel id
302.AIglm-5.3
AIHubMixglm-5.3
AgentRouterglm-5.3
Alibaba (China)glm-5.3
Alibaba Token Planglm-5.3
Alibaba Token Plan (China)glm-5.3
Basetenglm-5.3
Bothubglm-5.3
Charm Hyperglm-5.3
ClinePassglm-5.3
Cortecsglm-5.3
CrofAIglm-5.3

57 hosts serve this model at the price above · the maker's documentation

GLM: every version

VersionReleasedContextInput
GLM-5.3Current2026-08-141M$0.40
GLM-5.22026-06-131M$0.30
GLM-5.12026-04-07200K$0.45
GLM-5V-Turbo2026-04-01200K$0.7042
GLM-5-Turbo2026-03-16200K$0.72
GLM-52026-02-12205K$0.50

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

How large is the context window?
The context window is 1,000,000 tokens, and the maximum output is 131,072 tokens.
Is the model open weights?
Yes, the weights are open.
What is the pricing?
Input from $0.40 per million tokens, output from $1.40 per million tokens, cached input from $0.06 per million tokens.
Open weightsReasoning1M context