Skip to main content

Claude Sonnet 4.6 (Non-reasoning, High Effort)

AnthropicClaudeProprietary

Description

Claude Sonnet 4.6 is a full upgrade of the model's skills across coding, computer use, long-context reasoning, agent planning, knowledge work, and design. Users preferred Sonnet 4.6 over Sonnet 4.5 approximately 70% of the time. First Sonnet-class model with 1M token context window (beta) and context compaction. Major improvement in computer use skills compared to prior Sonnet models. Default model on Free and Pro plans. Pricing: $3/$15 per million tokens (input/output).

Release Date
2026-02-17
Parameters
Context Length
1.0M
Modalities
audio, image, pdf, text, video

Capability Radar

32
general
47
coding
80
reasoning
55
science
80
agents
80
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability39
53.0
LS
Code Ranking100
73.0
AA
General Ranking176
59.0
AA
Science161
61.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

BrowseCompOpenAI (2025)74.7%SR
OSWorld72.5%SR
Finance Agent63.3%SR
MCP Atlas61.3%SR
Terminal-Bench 2.0Stanford × Laude Institute (2026)59.1%SR
Finance Agent v251.0%
DeepSWE 1.130.0%
Legal Agent Benchmark5.4%

Biology

GPQANYU + Cohere + Anthropic (2023)89.9%SR

Code

SWE-Bench Verified79.6%SR

Communication

Tau2 Telecom97.9%SR
Tau2 Retail91.7%SR

General

MMMLU89.3%SR
MMMU-Pro75.6%SR
LiveBench75.5%

Math

Humanity's Last Exam49.0%SR

Reasoning

ARC-AGI v258.3%SR

AA Evaluation Indices

(Artificial Analysis)
Intelligence Index(Artificial Analysis)
36.8
Gpqa(NYU + Cohere + Anthropic (2023))
0.8
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.8
Lcr(Artificial Analysis)
0.6
Scicode(UIUC + Argonne National Lab (2024))
0.5
Terminalbench Hard(Stanford × Laude Institute (2026))
0.5
Ifbench(Google Research (2023))
0.4
Hle(Center for AI Safety + Scale AI (2025))
0.1

LLM Stats Category Scores

(LLM Stats (zeroeval))
Physics
90
Language
90
Biology
90
Chemistry
90
Communication
90
Frontend Development
80
Tool Calling
80
Math
70
Multimodal
70
Search
70
Reasoning
60
Spatial Reasoning
60
Finance
60
General
60
Code
60
Vision
60
Agents
50
Long Context
40
Healthcare
20
Legal
10

Pricing

Input Price$3 / 1M tokens
Output Price$15 / 1M tokens
Blended Price (3:1)$6 / 1M tokens
Cache Read Price$0.3 / 1M tokens
Cache Write Price$3.75 / 1M tokens

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

1 providers

ProviderInputOutput
1Anthropic
$0
$0.00002

Compare pricing across different API providers for this model.

External Sources