GLM-5.2 (max)
Описание
GLM-5.2 is Z.AI's flagship foundation model built for long-horizon tasks, delivering a solid 1M-token context that stably sustains long, messy coding-agent trajectories. It improves substantially over GLM-5.1, becoming the strongest open-source model on standard coding benchmarks (81.0 on Terminal-Bench 2.1 and 62.1 on SWE-bench Pro) and the highest-ranked open-source model across long-horizon coding benchmarks (FrontierSWE, PostTrainBench, SWE-Marathon). It introduces flexible thinking effort levels (High and Max) to balance capability against latency and compute. Architecturally, GLM-5.2 proposes IndexShare, which reuses one lightweight indexer across every four sparse-attention (DSA) layers to cut per-token FLOPs by 2.9x at 1M context, and an improved MTP layer for speculative decoding that raises acceptance length by up to 20%. Released under a pure MIT open-source license with weights available on HuggingFace and ModelScope, it supports transformers, vLLM, SGLang, xLLM, and ktransformers, with 1M input context, 128K max output, thinking mode, function calling, structured output, context caching, and MCP integration.
Радар способностей
Рейтинги
| Домен | #Место | Оценка | Источник |
|---|---|---|---|
| Агентные возможности | 34 | 51.0 | LS |
| Рейтинг кодинга | 72 | 83.0 | AA |
| Общий рейтинг | 32 | 78.0 | AA |
| Наука | 61 | 80.0 | AA |
Оценки бенчмарков (LLM Stats)
(LLM Stats (zeroeval))Agents
Code
Math
Reasoning
Индексы оценки AA
(Artificial Analysis)Оценки категорий LLM Stats
(LLM Stats (zeroeval))Цены
Скорость
Рейтинг цен провайдеров
Рейтинг цен провайдеров
9 провайдеров
Сравнение цен разных API-провайдеров для этой модели.