メインコンテンツへスキップ

GLM-5.1 (Reasoning)

Z AIGLMオープンウエイトMIT · 商用利用可

説明

GLM-5.1 is Z.AI's next-generation flagship foundation model designed for long-horizon agentic engineering tasks. Built on a 754B MoE architecture (40B active parameters), it can work continuously and autonomously on a single task for up to 8 hours, completing the full loop from planning and execution to iterative optimization and delivery. GLM-5.1 achieves state-of-the-art on SWE-Bench Pro (58.4) and demonstrates strong performance across coding, reasoning, and agentic benchmarks. It supports 200K context length, 128K max output tokens, thinking mode, function calling, structured output, context caching, and MCP integration. Overall performance is aligned with Claude Opus 4.6 with particular strengths in sustained execution and complex engineering optimization.

リリース日
2026-04-07
パラメータ
754.0B
コンテキスト長
200K
モダリティ
text

能力レーダー

39
general
54
coding
87
reasoning
60
science
60
agents
0
multimodal

ランキング

ドメイン#順位スコアソース
エージェント能力71
46.0
LS
コーディングランキング96
73.0
AA
総合ランキング54
79.0
AA
科学89
72.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Agents

Vending-Bench 2563441.0%自己申告
BrowseCompOpenAI (2025)79.3%自己申告
MCP Atlas71.8%自己申告
TAU3-Bench70.6%自己申告
Terminal-Bench 2.0Stanford × Laude Institute (2026)69.0%自己申告
CyberGym68.7%自己申告
SWE-Bench ProPrinceton NLP (2024)58.4%自己申告
Finance Agent v244.8%
NL2Repo42.7%自己申告
Toolathlon40.7%自己申告
FrontierSWE31.0%

Biology

GPQANYU + Cohere + Anthropic (2023)86.2%自己申告

General

LiveBench70.2%

Math

AIME 202695.3%自己申告
HMMT 202594.0%自己申告
IMO-AnswerBench83.8%自己申告
HMMT Feb 2682.6%自己申告
Humanity's Last Exam52.3%自己申告

AA評価指数

(Artificial Analysis)
Coding Index(Artificial Analysis)
55.8
Intelligence Index(Artificial Analysis)
41.0
Tau2(Sierra + U Toronto + Vector Institute (2025))
1.0
Gpqa(NYU + Cohere + Anthropic (2023))
0.9
Ifbench(Google Research (2023))
0.8
Lcr(Artificial Analysis)
0.7
Terminalbench V2 1
0.6
Scicode(UIUC + Argonne National Lab (2024))
0.4
Terminalbench Hard(Stanford × Laude Institute (2026))
0.4
Hle(Center for AI Safety + Scale AI (2025))
0.3
Tau Banking
0.1

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Agents
100
Reasoning
100
General
100
Physics
90
Biology
90
Chemistry
90
Math
80
Search
80
Safety
70
Code
60
Tool Calling
60
Vision
50
Finance
40

価格設定

入力価格$1.38 / 1Mトークン
出力価格$4.4 / 1Mトークン
混合価格(3:1)$2.135 / 1Mトークン
キャッシュ読み取り価格$0.26 / 1Mトークン
キャッシュ書き込み価格無料

速度

トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s

プロバイダー価格ランキング

プロバイダー価格ランキング

20 プロバイダー

最安: ZAI最高: GreenPT
プロバイダー入力出力
1ZAI最安
$0
$0
2FriendliAI
$0
$0
3CrofAI
$0.45
$2.15
4EmpirioLabs AI
$0.825
$3.301
5EBCloud
$0.8571
$3.4286
6302.AI
$0.86
$3.5
7Alibaba (China)
$0.87
$3.48
8LLM Gateway
$0.931
$2.93
9DigitalOcean
$0.975
$4.3
10Wafer
$1
$3.2
11DInference
$1.25
$3.89
12Z AIプライマリ
$1.38
$4.4
13Cortecs
$1.384
$4.348
14OpenCode Go
$1.4
$4.4
15Z.AI
$1.4
$4.4
16OpenCode Zen
$1.4
$4.4
17Zhipu AI
$1.4
$4.4
18Auriko
$1.4
$4.4
19Charm Hyper
$1.52432
$4.79072
20GreenPT
$1.756
$5.518

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク