Claude 3.7 Sonnet (Non-reasoning)
AnthropicClaudeProprietary
설명
The most intelligent Claude model and the first hybrid reasoning model on the market. Claude 3.7 Sonnet can produce near-instant responses or extended, step-by-step thinking that is made visible to the user. Shows particularly strong improvements in coding and front-end web development.
출시일
2025-02-24
파라미터
—
컨텍스트 길이
—
모달리티
image, text
능력 레이더
33
general
39
coding
35
reasoning
47
science
70
agents
80
multimodal
랭킹
벤치마크 점수 (LLM Stats)
(LLM Stats (zeroeval))Chat
IFEvalGoogle Research (2023)
93.2%자체 보고
TAU-bench Retail
81.2%자체 보고
Language
MMMLU
86.1%자체 보고
Math
MATH-500
96.2%자체 보고
AIME 2024
80.0%자체 보고
AIME 2025
54.8%자체 보고
Multimodal
MMMU
75.0%자체 보고
Reasoning
GPQANYU + Cohere + Anthropic (2023)
84.8%자체 보고
SWE-Bench Verified
70.3%자체 보고
TAU-bench Airline
58.4%자체 보고
Terminal-Bench
35.2%자체 보고
AA 평가 지수
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))85.0
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))80.3
Gpqa(NYU + Cohere + Anthropic (2023))65.6
Lcr(Artificial Analysis)51.7
Tau2(Sierra + U Toronto + Vector Institute (2025))50.0
Ifbench(Google Research (2023))44.0
Livecodebench(UC Berkeley + MIT + Cornell (2024))39.4
Aime(MAA (Mathematical Association of America))22.3
Terminalbench Hard(Stanford × Laude Institute (2026))21.2
Math Index(Artificial Analysis)21.0
Aime 25(MAA (Mathematical Association of America))21.0
Intelligence Index(Artificial Analysis)15.3
Hle(Center for AI Safety + Scale AI (2025))4.2
LLM Stats 카테고리 점수
(LLM Stats (zeroeval))Chat90
Instruction Following90
Language90
Structured Output90
Math80
Multimodal80
Physics80
Healthcare80
Biology80
Chemistry80
Vision80
Reasoning70
Frontend Development70
General70
Communication70
Tool Calling70
Code50
Agents40
가격
입력 가격$3 / 1M 토큰
출력 가격$15 / 1M 토큰
혼합 가격 (3:1)$6 / 1M 토큰
속도
토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s
공급자 가격 순위
공급자 가격 순위
1개 공급자
공급자입력출력
1Anthropic주요
$3
$15
이 모델의 다양한 API 공급자 간 가격 비교.