Skip to main content

Claude Sonnet 5.5 (Adaptive Reasoning, Max Effort, Default Fallback)

AnthropicClaudeProprietary

Description

Claude Sonnet 5.5 is the second model in Anthropic's Claude 5.5 family and a faster, lower-cost complement to Claude Opus 5.5. It targets well-scoped everyday tasks, coding, and polished documents, slides, and spreadsheets, and runs 30%+ faster than Claude Sonnet 5 while typically using far fewer tokens at the same $2/$10 per million input/output prices. It accepts text and image input with text output, supports multilingual and vision workloads, tool use, and adaptive thinking (Claude Platform default effort high; Claude Code and apps default medium). Context window is 1M tokens with 128K max output on the sync Messages API; reliable knowledge and training-data cutoffs are June 2026. Cache reads are $0.20 and 5-minute cache writes are $2.50 per million tokens. Capability scores are self-reported from the Claude Sonnet 5.5 System Card (Table 8.1.A and §8) and the launch announcement. Available on the Claude API as `claude-sonnet-5-5`.

Release Date
2026-09-28
Parameters
—
Context Length
1.0M
Modalities
image, pdf, text

Capability Radar

56
general
61
coding
70
reasoning
59
science
60
agents
80
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability14
63.0
LS
Code Ranking31
93.0
AA
General Ranking3
97.0
AA
Science9
90.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

AA-Briefcase v1.11811.00 / 3000SR
Program Bench79.7%SR
Toolathlon-Verified77.8%SR
OfficeQA76.9%SR
OfficeQA Pro65.6%SR
AutomationBench44.7%SR

Biology

BioMysteryBench89.2%SR
De novo protein binder design82.3%SR
LatchBio SpatialBench Verified72.5%SR
Biomedical image analysis72.2%SR
Protocols Troubleshooting67.3%SR
Protocols Understanding (V2)66.6%SR
Medicinal Chemistry (ADME)65.3%SR
Protein Design — Library Ranking54.8%SR
Protein Design — Sequence Generation51.0%SR
BioMysteryBench (Human Difficult)44.7%SR
Morphology-to-molecule matching25.0%SR

Code

BenchCAD (with Python tool)96.3%SR
BenchCAD74.7%SR
DeepSWE 1.171.0%SR
FrontierSWE V261.9%SR
CursorBench 4.055.5%SR

General

GDPval-AA 2.11844.00 / 3000SR
Global-MMLU92.1%SR

Healthcare

HealthBench Professional (raw)77.1%SR
HealthBench (raw)69.4%SR
HealthBench Professional69.2%SR
HealthBench65.4%SR
PhysicianBench63.2%SR

Language

MILU91.6%SR

Legal

Legal Agent Benchmark11.7%SR

Math

ArXivMath (with tools)95.2%SR
ArXivMath86.8%SR

Multimodal

OSWorld 2.1 (partial)80.1%SR
OSWorld 2.1 (strict)43.5%SR

Reasoning

SWE-bench Multilingual90.3%SR
SWE-Bench ProPrinceton NLP (2024)81.3%SR
Terminal-Bench 4.070.6%SR
Humanity's Last Exam (with tools)64.5%SR
FrontierCode 1.1 (Extended)64.4%SR
Terminal-Bench-Science 0.159.9%SR
Humanity's Last Exam (no tools, text-only)56.9%SR
SWE-Bench Multimodal54.3%SR
FrontierCode 1.152.1%SR

Science

LatchBio SingleCellBench59.1%SR

Vision

Chartography90.2%SR
Chartography (no tools)61.6%SR

AA Evaluation Indices

(Artificial Analysis)
Lcr(Artificial Analysis)
82.7
Scicode(UIUC + Argonne National Lab (2024))
61.0
Intelligence Index(Artificial Analysis)
56.0
Hle(Center for AI Safety + Scale AI (2025))
55.0

LLM Stats Category Scores

(LLM Stats (zeroeval))
Productivity
100
Agents
100
Reasoning
100
General
100
Language
90
Science
90
Biology
90
Multimodal
80
Vision
80
Math
70
Healthcare
70
Code
70
Tool Calling
60
Legal
10

Pricing

Input Price$2 / 1M tokens
Output Price$10 / 1M tokens
Blended Price (3:1)$4 / 1M tokens
Cache Read Price$0.2 / 1M tokens
Cache Write Price$2.5 / 1M tokens

Speed

Tokens/sec145.1
Time to First Token347.54s
Time to Answer347.54s

Provider Price Ranking

Provider Price Ranking

1 providers

ProviderInputOutput
1Anthropic
$0
$0.00001

Compare pricing across different API providers for this model.

External Sources