Skip to main content

GLM-4.5V (Non-reasoning)

Z AIGLMOpen WeightMIT · Commercial OK

Description

GLM-4.5V is a multimodal (vision-language) model based on GLM-4.5-Air (106B total, 12B active) that extends hybrid reasoning to images and video. It achieves state-of-the-art results across 40+ VLM benchmarks (image reasoning, video understanding, GUI tasks, chart/document parsing, grounding) while supporting a Thinking Mode switch for deep reasoning. Released under MIT with FP8/BF16 variants and tooling in Transformers, vLLM, and SGLang.

Release Date
2025-08-11
Parameters
108.0B
Context Length
64K
Modalities
image, text, video

Capability Radar

27
general
32
coding
22
reasoning
33
science
25
agents
90
multimodal

Rankings

Domain#RankScoreSource
Code Ranking462
16.0
AA
General Ranking447
29.0
AA
Science439
30.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

No benchmark data available

AA Evaluation Indices

(Artificial Analysis)
Math Index(Artificial Analysis)
15.3
Intelligence Index(Artificial Analysis)
6.8
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.8
Gpqa(NYU + Cohere + Anthropic (2023))
0.6
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.4
Ifbench(Google Research (2023))
0.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.2
Scicode(UIUC + Argonne National Lab (2024))
0.2
Aime 25(MAA (Mathematical Association of America))
0.2
Terminalbench Hard(Stanford × Laude Institute (2026))
0.1
Hle(Center for AI Safety + Scale AI (2025))
0.0
Lcr(Artificial Analysis)
0.0

LLM Stats Category Scores

(LLM Stats (zeroeval))

No category score data available

Pricing

Input Price$0.6 / 1M tokens
Output Price$1.8 / 1M tokens
Blended Price (3:1)$0.9 / 1M tokens

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

1 providers

ProviderInputOutput
1Z AIPRIMARY
$0.6
$1.8

Compare pricing across different API providers for this model.

External Sources