메인 콘텐츠로 건너뛰기

GLM-5.1 (Reasoning)

Z AIGLM오픈 웨이트MIT · 상업적 사용 가능

설명

GLM-5.1 is Z.AI's next-generation flagship foundation model designed for long-horizon agentic engineering tasks. Built on a 754B MoE architecture (40B active parameters), it can work continuously and autonomously on a single task for up to 8 hours, completing the full loop from planning and execution to iterative optimization and delivery. GLM-5.1 achieves state-of-the-art on SWE-Bench Pro (58.4) and demonstrates strong performance across coding, reasoning, and agentic benchmarks. It supports 200K context length, 128K max output tokens, thinking mode, function calling, structured output, context caching, and MCP integration. Overall performance is aligned with Claude Opus 4.6 with particular strengths in sustained execution and complex engineering optimization.

출시일
2026-04-07
파라미터
754.0B
컨텍스트 길이
200K
모달리티
text

능력 레이더

27
general
54
coding
87
reasoning
61
science
60
agents
0
multimodal

랭킹

도메인#순위점수소스
에이전트형 역량44
49.0
LS
코딩 랭킹168
71.0
AA
종합 랭킹90
68.0
AA
과학143
67.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Agents

TAU3-Bench70.6%자체 보고
Finance Agent v244.8%
Toolathlon40.7%자체 보고

Code

CyberGym68.7%자체 보고
NL2Repo42.7%자체 보고
FrontierSWE31.0%

Math

AIME 202695.3%자체 보고
HMMT 202594.0%자체 보고
IMO-AnswerBench83.8%자체 보고
HMMT Feb 2682.6%자체 보고
LiveBench70.2%

Reasoning

Vending-Bench 2563441.0%자체 보고
GPQANYU + Cohere + Anthropic (2023)86.2%자체 보고
BrowseCompOpenAI (2025)79.3%자체 보고
MCP Atlas71.8%자체 보고
Terminal-Bench 2.0Stanford × Laude Institute (2026)69.0%자체 보고
SWE-Bench ProPrinceton NLP (2024)58.4%자체 보고
Humanity's Last Exam52.3%자체 보고

AA 평가 지수

(Artificial Analysis)
Tau2(Sierra + U Toronto + Vector Institute (2025))
97.7
Gpqa(NYU + Cohere + Anthropic (2023))
86.8
Ifbench(Google Research (2023))
76.3
Lcr(Artificial Analysis)
73.7
Terminalbench V2 1
61.8
Coding Index(Artificial Analysis)
55.8
Scicode(UIUC + Argonne National Lab (2024))
44.8
Terminalbench Hard(Stanford × Laude Institute (2026))
43.2
Hle(Center for AI Safety + Scale AI (2025))
30.1
Intelligence Index(Artificial Analysis)
26.1
Tau Banking
13.6
Terminalbench V4 0
2.0

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Agents
100
Reasoning
100
General
100
Physics
90
Biology
90
Chemistry
90
Math
80
Search
80
Safety
70
Code
60
Tool Calling
60
Vision
50
Finance
40

가격

입력 가격$1.285 / 1M 토큰
출력 가격$4.07 / 1M 토큰
혼합 가격 (3:1)$1.981 / 1M 토큰
캐시 읽기 가격$0.26 / 1M 토큰
캐시 쓰기 가격무료

속도

토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s

공급자 가격 순위

공급자 가격 순위

5개 공급자

최저가: DeepInfra최고가: Z AI
공급자입력출력
1DeepInfra최저가
$0
$0
2ZAI
$0
$0
3FriendliAI
$0
$0
4EmpirioLabs AI
$0.825
$3.301
5Z AI주요
$1.285
$4.07

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크