GLM-4.6 (Reasoning)
Z AIGLM開源權重MIT · 商用許可
描述
GLM-4.6 is the latest version of Z.ai's flagship model, bringing significant improvements over GLM-4.5. Key features include: 200K token context window (expanded from 128K), superior coding performance with better real-world application in Claude Code/Cline/Roo Code/Kilo Code, advanced reasoning with tool use during inference, stronger agent capabilities, and refined writing aligned with human preferences. GLM-4.6 achieves competitive performance with DeepSeek-V3.2-Exp and Claude Sonnet 4, reaching near parity with Claude Sonnet 4 (48.6% win rate) on CC-Bench real-world coding tasks.
發布日期
2025-09-30
參數規模
357.0B
上下文長度
205K
支援模態
image, text, video
能力雷達圖
37
general
55
coding
85
reasoning
58
science
40
agents
20
multimodal
排行榜排名
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))Math
AIME 2025
93.9%自報
Reasoning
LiveCodeBench v6
82.8%自報
GPQANYU + Cohere + Anthropic (2023)
81.0%自報
SWE-Bench Verified
68.0%自報
BrowseCompOpenAI (2025)
45.1%自報
Terminal-Bench
40.5%自報
Humanity's Last Exam
17.2%自報
AA 評測指數
(Artificial Analysis)Math Index(Artificial Analysis)86.0
Aime 25(MAA (Mathematical Association of America))86.0
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))82.9
Gpqa(NYU + Cohere + Anthropic (2023))78.0
Tau2(Sierra + U Toronto + Vector Institute (2025))70.5
Livecodebench(UC Berkeley + MIT + Cornell (2024))69.5
Lcr(Artificial Analysis)54.0
Terminalbench V2 149.4
Coding Index(Artificial Analysis)45.8
Ifbench(Google Research (2023))43.4
Terminalbench Hard(Stanford × Laude Institute (2026))25.0
Intelligence Index(Artificial Analysis)18.5
Hle(Center for AI Safety + Scale AI (2025))14.5
Tau Banking13.4
LLM Stats 分類評分
(LLM Stats (zeroeval))Physics80
Biology80
Chemistry80
Frontend Development70
Math60
Reasoning60
General60
Search50
Code50
Agents40
Vision20
定價
輸入價格$0.55 / 1M tokens
輸出價格$2.2 / 1M tokens
混合價格(3:1)$0.963 / 1M tokens
快取讀取價格$0.11 / 1M tokens
快取寫入價格免費
速度
Tokens/秒0.0
首Token延遲0.00s
首回答延遲0.00s
供應商價格排行
供應商價格排行
2 個供應商
最便宜: DeepInfra最貴: Z AI
供應商輸入輸出
1DeepInfra最便宜
$0
$0
2Z AI主要
$0.55
$2.2
比較該模型在不同 API 供應商之間的定價。