25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / Z.ai (Zhipu, GLM) / GLM Air / GLM-4.5-Air

GLM-4.5-Air

Lighter GLM-4.5 variant for fast coding assistance and cheaper agents

Line

GLM Air

Weights

Open

Released

2025-07-28

Context

131K

Input

$0.10 / 1M

Output

$0.2911 / 1M

Max output

98K

Coverage

not covered yet

Overview

GLM-4.5-Air is a lighter variant of GLM-4.5 from Z.ai (Zhipu, GLM). It is designed for fast coding assistance and cheaper agents. It is the first version of the glm-air line. It has a context window of 131072 tokens and a maximum output of 98304 tokens. Weights are open. It supports reasoning mode and tool calling. It accepts text and returns text. Knowledge cutoff is April 2025. It is served by 19 hosts. Pricing starts at $0.10 per million input tokens and $0.2911 per million output tokens. Cached input is $0.0233 per million tokens.

Specs

Released2025-07-28
LineZ.ai (Zhipu, GLM) · GLM Air
WeightsOpen
Context131K tokens
Max output98K tokens
Inputtext
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoff2025-04
API idglm-4.5-air

Pricing

Input$0.10 / 1M tokens
Cached input$0.0233 / 1M tokens
Output$0.2911 / 1M tokens

First version of its line in the registry.

Range across the hosts that serve it: up to $0.20 per 1M input tokens.

Strengths

  • 131072 token context window
  • 98304 token max output
  • Open weights
  • Reasoning mode and tool calling
  • Priced from $0.10 per million input

Best for

  • Reach for it for fast coding assistance
  • Reach for it for cheaper agent workflows
  • Reach for it for text generation with tool calling
  • Reach for it for reasoning tasks

How to access

ProviderModel id
DevPass (LLM Gateway)glm-4.5-air
Hugging Faceglm-4.5-air
Impossiblglm-4.5-air
Kilo Gatewayglm-4.5-air
LLM Gatewayglm-4.5-air
Merge Gatewayglm-4.5-air
NanoGPTglm-4.5-air
NovitaAIglm-4.5-air
OpenRouterglm-4.5-air
OrcaRouterglm-4.5-air
Qiniuglm-4.5-air
SiliconFlowglm-4.5-air

19 hosts serve this model at the price above · the maker's documentation

GLM Air: every version

VersionReleasedContextInput
GLM-4.5-AirCurrent2025-07-28131K$0.10

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window?
The context window is 131072 tokens, and the maximum output is 98304 tokens.
Are the weights open?
Yes, the weights are open. The model also supports reasoning mode and tool calling.
What is the price?
Input starts at $0.10 per million tokens, output at $0.2911 per million, and cached input at $0.0233 per million.
Open weightsReasoning131K context