メインコンテンツへスキップ

GLM-5.2 (max)

Z AIGLMオープンウエイトMIT · 商用利用可

説明

GLM-5.2 is Z.AI's flagship foundation model built for long-horizon tasks, delivering a solid 1M-token context that stably sustains long, messy coding-agent trajectories. It improves substantially over GLM-5.1, becoming the strongest open-source model on standard coding benchmarks (81.0 on Terminal-Bench 2.1 and 62.1 on SWE-bench Pro) and the highest-ranked open-source model across long-horizon coding benchmarks (FrontierSWE, PostTrainBench, SWE-Marathon). It introduces flexible thinking effort levels (High and Max) to balance capability against latency and compute. Architecturally, GLM-5.2 proposes IndexShare, which reuses one lightweight indexer across every four sparse-attention (DSA) layers to cut per-token FLOPs by 2.9x at 1M context, and an improved MTP layer for speculative decoding that raises acceptance length by up to 20%. Released under a pure MIT open-source license with weights available on HuggingFace and ModelScope, it supports transformers, vLLM, SGLang, xLLM, and ktransformers, with 1M input context, 128K max output, thinking mode, function calling, structured output, context caching, and MCP integration.

リリース日
2026-06-16
パラメータ
753.0B
コンテキスト長
1.0M
モダリティ
text

能力レーダー

39
general
66
coding
90
reasoning
66
science
70
agents
0
multimodal

ランキング

ドメイン#順位スコアソース
エージェント能力34
51.0
LS
コーディングランキング71
83.0
AA
総合ランキング25
82.0
AA
科学61
79.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Agents

Program Bench63.7%自己申告
Toolathlon48.2%自己申告
PostTrainBench34.3%自己申告

Code

FrontierSWE74.0%
NL2Repo48.9%自己申告
DeepSWE46.2%自己申告
DeepSWE 1.144.0%
SWE-Marathon13.0%自己申告

Math

AIME 202699.2%自己申告
HMMT 202594.4%自己申告
HMMT Feb 2692.5%自己申告
IMO-AnswerBench91.0%自己申告

Reasoning

GPQANYU + Cohere + Anthropic (2023)91.2%自己申告
Terminal-Bench 2.182.7%自己申告
MCP Atlas76.8%自己申告
SWE-Bench ProPrinceton NLP (2024)62.1%自己申告
Humanity's Last Exam54.7%自己申告
FrontierCode 1.124.5%
CritPT16.7%自己申告

AA評価指数

(Artificial Analysis)
Tau2(Sierra + U Toronto + Vector Institute (2025))
99.1
Gpqa(NYU + Cohere + Anthropic (2023))
89.5
Lcr(Artificial Analysis)
78.3
Terminalbench V2 1
77.9
Ifbench(Google Research (2023))
73.3
Coding Index(Artificial Analysis)
68.8
Scicode(UIUC + Argonne National Lab (2024))
51.2
Terminalbench Hard(Stanford × Laude Institute (2026))
50.8
Hle(Center for AI Safety + Scale AI (2025))
41.1
Intelligence Index(Artificial Analysis)
38.6
Tau Banking
34.6

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Physics
90
Biology
90
Chemistry
90
Math
70
Tool Calling
70
Reasoning
60
General
60
Agents
50
Code
50
Vision
50
Systems
30

価格設定

入力価格$1.4 / 1Mトークン
出力価格$4.4 / 1Mトークン
混合価格(3:1)$2.15 / 1Mトークン
キャッシュ読み取り価格$0.26 / 1Mトークン
キャッシュ書き込み価格無料

速度

トークン/秒71.9
初トークン遅延7.81s
初回答遅延35.61s

プロバイダー価格ランキング

プロバイダー価格ランキング

9 プロバイダー

最安: DeepInfra最高: EmpirioLabs AI
プロバイダー入力出力
1DeepInfra最安
$0
$0
2Fireworks
$0
$0
3Novita
$0
$0
4FriendliAI
$0
$0
5Together
$0
$0
6ZAI
$0
$0
7Z AIプライマリ
$1.4
$4.4
8Neon
$1.4
$4.4
9EmpirioLabs AI
$1.4
$4.4

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク