Skip to main content

GLM-4.5-Air

Z AIGLMOpen WeightMIT · Commercial OK

Description

GLM-4.5-Air is a more compact variant of GLM-4.5 designed for efficient Agentic, Reasoning, and Coding (ARC) applications. It features 106 billion total parameters with 12 billion active parameters using MoE architecture. Like GLM-4.5, it is a hybrid reasoning model providing thinking mode for complex reasoning and tool usage, and non-thinking mode for immediate responses. Despite its compact design, GLM-4.5-Air delivers competitive performance with a score of 59.8 across 12 industry-standard benchmarks, ranking 6th overall while maintaining superior efficiency. It supports 128K context length and is released under MIT open-source license allowing commercial use.

Release Date
2025-07-28
Parameters
106.0B
Context Length
131K
Modalities
text

Capability Radar

32
general
68
coding
79
reasoning
53
science
70
agents
0
multimodal

Rankings

Domain#RankScoreSource
Code Ranking287
53.0
AA
General Ranking333
41.0
AA
Science352
42.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Chat

TAU-bench Retail77.9%SR

General

BFCL-v376.4%SR
AA-Index64.8%SR

Language

MMLU-Pro81.4%SR

Math

MATH-50098.1%SR
AIME 202489.4%SR

Reasoning

GPQANYU + Cohere + Anthropic (2023)75.0%SR
LiveCodeBench70.7%SR
TAU-bench Airline60.8%SR
SWE-Bench Verified57.6%SR
SciCode37.3%SR
Terminal-Bench30.0%SR
BrowseCompOpenAI (2025)21.3%SR
Humanity's Last Exam10.6%SR

AA Evaluation Indices

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
96.5
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
81.5
Math Index(Artificial Analysis)
80.7
Aime 25(MAA (Mathematical Association of America))
80.7
Gpqa(NYU + Cohere + Anthropic (2023))
73.3
Livecodebench(UC Berkeley + MIT + Cornell (2024))
68.4
Aime(MAA (Mathematical Association of America))
67.3
Lcr(Artificial Analysis)
46.7
Tau2(Sierra + U Toronto + Vector Institute (2025))
46.5
Ifbench(Google Research (2023))
37.6
Terminalbench Hard(Stanford × Laude Institute (2026))
20.5
Intelligence Index(Artificial Analysis)
11.1
Hle(Center for AI Safety + Scale AI (2025))
7.0

LLM Stats Category Scores

(LLM Stats (zeroeval))
Chat
80
Language
80
Legal
80
Structured Output
80
Finance
80
Healthcare
80
Communication
70
Tool Calling
70
Math
60
Physics
60
Reasoning
60
Frontend Development
60
General
60
Biology
60
Chemistry
60
Code
50
Agents
40
Search
20
Vision
10

Pricing

Input Price$0.17 / 1M tokens
Output Price$0.98 / 1M tokens
Blended Price (3:1)$0.372 / 1M tokens
Cache Read Price$0.03 / 1M tokens
Cache Write PriceFree

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

8 providers

Cheapest: ZenMuxMost Expensive: OrcaRouter
ProviderInputOutput
1ZenMuxCheapest
$0.1165
$0.2911
2OpenRouter
$0.13
$0.85
3Kilo Gateway
$0.13
$0.85
4DevPass (LLM Gateway)
$0.13
$0.85
5Z AIPRIMARY
$0.17
$0.98
6Z.AI
$0.2
$1.1
7Zhipu AI
$0.2
$1.1
8OrcaRouter
$0.2
$1.1

Compare pricing across different API providers for this model.

External Sources