Skip to main content

Claude 4 Sonnet (Reasoning)

AnthropicClaude
Release Date
2025-05-22
Parameters
—
Context Length
—
Modalities
—

Capability Radar

37
general
48
coding
79
reasoning
57
science
70
agents
80
multimodal

Rankings

Domain#RankScoreSource
Code Ranking252
58.0
AA
General Ranking195
55.0
AA
Science295
47.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

No benchmark data available

AA Evaluation Indices

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
99.1
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
84.2
Gpqa(NYU + Cohere + Anthropic (2023))
77.7
Aime(MAA (Mathematical Association of America))
77.3
Aime 25(MAA (Mathematical Association of America))
74.3
Math Index(Artificial Analysis)
74.3
Lcr(Artificial Analysis)
70.3
Livecodebench(UC Berkeley + MIT + Cornell (2024))
65.5
Tau2(Sierra + U Toronto + Vector Institute (2025))
64.6
Ifbench(Google Research (2023))
54.7
Coding Index(Artificial Analysis)
37.6
Terminalbench V2 1
36.3
Terminalbench Hard(Stanford × Laude Institute (2026))
31.1
Intelligence Index(Artificial Analysis)
18.9
Tau Banking
16.7
Hle(Center for AI Safety + Scale AI (2025))
10.7

LLM Stats Category Scores

(LLM Stats (zeroeval))

No category score data available

Pricing

Input PriceFree
Output PriceFree
Blended Price (3:1)Free

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

No provider data available

External Sources