GLM-4.5V (Non-reasoning)
Z AIGLMOpen WeightMIT · Commercial OK
Description
GLM-4.5V is a multimodal (vision-language) model based on GLM-4.5-Air (106B total, 12B active) that extends hybrid reasoning to images and video. It achieves state-of-the-art results across 40+ VLM benchmarks (image reasoning, video understanding, GUI tasks, chart/document parsing, grounding) while supporting a Thinking Mode switch for deep reasoning. Released under MIT with FP8/BF16 variants and tooling in Transformers, vLLM, and SGLang.
Release Date
2025-08-11
Parameters
108.0B
Context Length
64K
Modalities
image, text, video
Capability Radar
27
general
32
coding
22
reasoning
33
science
25
agents
90
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 462 | 16.0 | AA |
| General Ranking | 447 | 29.0 | AA |
| Science | 439 | 30.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))No benchmark data available
AA Evaluation Indices
(Artificial Analysis)Math Index(Artificial Analysis)15.3
Intelligence Index(Artificial Analysis)6.8
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.8
Gpqa(NYU + Cohere + Anthropic (2023))0.6
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.4
Ifbench(Google Research (2023))0.3
Tau2(Sierra + U Toronto + Vector Institute (2025))0.2
Scicode(UIUC + Argonne National Lab (2024))0.2
Aime 25(MAA (Mathematical Association of America))0.2
Terminalbench Hard(Stanford × Laude Institute (2026))0.1
Hle(Center for AI Safety + Scale AI (2025))0.0
Lcr(Artificial Analysis)0.0
LLM Stats Category Scores
(LLM Stats (zeroeval))No category score data available
Pricing
Input Price$0.6 / 1M tokens
Output Price$1.8 / 1M tokens
Blended Price (3:1)$0.9 / 1M tokens
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
Provider Price Ranking
1 providers
ProviderInputOutput
1Z AIPRIMARY
$0.6
$1.8
Compare pricing across different API providers for this model.