GLM-4.6 (Reasoning)
Z AIGLMOpen WeightMIT · Commercial OK
Description
GLM-4.6 is the latest version of Z.ai's flagship model, bringing significant improvements over GLM-4.5. Key features include: 200K token context window (expanded from 128K), superior coding performance with better real-world application in Claude Code/Cline/Roo Code/Kilo Code, advanced reasoning with tool use during inference, stronger agent capabilities, and refined writing aligned with human preferences. GLM-4.6 achieves competitive performance with DeepSeek-V3.2-Exp and Claude Sonnet 4, reaching near parity with Claude Sonnet 4 (48.6% win rate) on CC-Bench real-world coding tasks.
Release Date
2025-09-30
Parameters
357.0B
Context Length
205K
Modalities
image, text, video
Capability Radar
37
general
55
coding
85
reasoning
58
science
40
agents
20
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 252 | 57.0 | AA |
| General Ranking | 207 | 53.0 | AA |
| Science | 256 | 51.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Math
AIME 2025
93.9%SR
Reasoning
LiveCodeBench v6
82.8%SR
GPQANYU + Cohere + Anthropic (2023)
81.0%SR
SWE-Bench Verified
68.0%SR
BrowseCompOpenAI (2025)
45.1%SR
Terminal-Bench
40.5%SR
Humanity's Last Exam
17.2%SR
AA Evaluation Indices
(Artificial Analysis)Math Index(Artificial Analysis)86.0
Aime 25(MAA (Mathematical Association of America))86.0
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))82.9
Gpqa(NYU + Cohere + Anthropic (2023))78.0
Tau2(Sierra + U Toronto + Vector Institute (2025))70.5
Livecodebench(UC Berkeley + MIT + Cornell (2024))69.5
Lcr(Artificial Analysis)54.0
Terminalbench V2 149.4
Coding Index(Artificial Analysis)45.8
Ifbench(Google Research (2023))43.4
Terminalbench Hard(Stanford × Laude Institute (2026))25.0
Intelligence Index(Artificial Analysis)18.5
Hle(Center for AI Safety + Scale AI (2025))14.5
Tau Banking13.4
LLM Stats Category Scores
(LLM Stats (zeroeval))Physics80
Biology80
Chemistry80
Frontend Development70
Math60
Reasoning60
General60
Search50
Code50
Agents40
Vision20
Pricing
Input Price$0.55 / 1M tokens
Output Price$2.2 / 1M tokens
Blended Price (3:1)$0.963 / 1M tokens
Cache Read Price$0.11 / 1M tokens
Cache Write PriceFree
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
Provider Price Ranking
2 providers
Cheapest: DeepInfraMost Expensive: Z AI
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2Z AIPRIMARY
$0.55
$2.2
Compare pricing across different API providers for this model.