GLM-4.7-Flash (Non-reasoning)
Z AIGLM오픈 웨이트MIT · 상업적 사용 가능
설명
GLM-4.7-Flash is a high-speed, cost-efficient variant of GLM-4.7 optimized for fast inference and lower latency. It retains the coding-centric capabilities of GLM-4.7 including thinking before acting, preserved reasoning across turns, and per-request thinking control for speed or accuracy trade-offs. Ideal for applications requiring quick responses while maintaining strong performance on coding, agentic workflows, and general reasoning tasks.
출시일
2026-01-19
파라미터
30.0B
컨텍스트 길이
200K
모달리티
text
능력 레이더
13
general
26
coding
45
reasoning
30
science
80
agents
0
multimodal
랭킹
벤치마크 점수 (LLM Stats)
(LLM Stats (zeroeval))Agents
Tau-bench
79.5%자체 보고
BrowseCompOpenAI (2025)
42.8%자체 보고
Biology
GPQANYU + Cohere + Anthropic (2023)
75.2%자체 보고
Code
SWE-Bench Verified
59.2%자체 보고
Math
AIME 2025
91.6%자체 보고
Humanity's Last Exam
14.4%자체 보고
AA 평가 지수
(Artificial Analysis)Intelligence Index(Artificial Analysis)15.6
Tau2(Sierra + U Toronto + Vector Institute (2025))0.9
Ifbench(Google Research (2023))0.5
Gpqa(NYU + Cohere + Anthropic (2023))0.5
Scicode(UIUC + Argonne National Lab (2024))0.3
Lcr(Artificial Analysis)0.2
Hle(Center for AI Safety + Scale AI (2025))0.1
Terminalbench Hard(Stanford × Laude Institute (2026))0.0
LLM Stats 카테고리 점수
(LLM Stats (zeroeval))Physics80
Biology80
Chemistry80
Tool Calling80
Reasoning60
Frontend Development60
General60
Agents60
Code60
Math50
Search40
Vision10
가격
입력 가격$0.07 / 1M 토큰
출력 가격$0.4 / 1M 토큰
혼합 가격 (3:1)$0.153 / 1M 토큰
캐시 읽기 가격무료
캐시 쓰기 가격무료
속도
토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s
공급자 가격 순위
공급자 가격 순위
1개 공급자
공급자입력출력
1Z AI주요
$0.07
$0.4
이 모델의 다양한 API 공급자 간 가격 비교.