跳转到主要内容

Claude Sonnet 5.5 (Adaptive Reasoning, Max Effort, Default Fallback)

AnthropicClaudeProprietary

描述

Claude Sonnet 5.5 is the second model in Anthropic's Claude 5.5 family and a faster, lower-cost complement to Claude Opus 5.5. It targets well-scoped everyday tasks, coding, and polished documents, slides, and spreadsheets, and runs 30%+ faster than Claude Sonnet 5 while typically using far fewer tokens at the same $2/$10 per million input/output prices. It accepts text and image input with text output, supports multilingual and vision workloads, tool use, and adaptive thinking (Claude Platform default effort high; Claude Code and apps default medium). Context window is 1M tokens with 128K max output on the sync Messages API; reliable knowledge and training-data cutoffs are June 2026. Cache reads are $0.20 and 5-minute cache writes are $2.50 per million tokens. Capability scores are self-reported from the Claude Sonnet 5.5 System Card (Table 8.1.A and §8) and the launch announcement. Available on the Claude API as `claude-sonnet-5-5`.

发布日期
2026-09-28
参数规模
—
上下文长度
1.0M
支持模态
image, pdf, text

能力雷达图

56
general
61
coding
70
reasoning
59
science
60
agents
80
multimodal

排行榜排名

领域#排名分数来源
智能体能力模型榜14
63.0
LS
代码能力榜34
93.0
AA
通用能力榜3
97.0
AA
科学能力9
90.0
AA

基准测试分数 (LLM Stats)

(LLM Stats (zeroeval))

Agents

AA-Briefcase v1.11811.00 / 3000自报
Program Bench79.7%自报
Toolathlon-Verified77.8%自报
OfficeQA76.9%自报
OfficeQA Pro65.6%自报
AutomationBench44.7%自报

Biology

BioMysteryBench89.2%自报
De novo protein binder design82.3%自报
LatchBio SpatialBench Verified72.5%自报
Biomedical image analysis72.2%自报
Protocols Troubleshooting67.3%自报
Protocols Understanding (V2)66.6%自报
Medicinal Chemistry (ADME)65.3%自报
Protein Design — Library Ranking54.8%自报
Protein Design — Sequence Generation51.0%自报
BioMysteryBench (Human Difficult)44.7%自报
Morphology-to-molecule matching25.0%自报

Code

BenchCAD (with Python tool)96.3%自报
BenchCAD74.7%自报
DeepSWE 1.171.0%自报
FrontierSWE V261.9%自报
CursorBench 4.055.5%自报

General

GDPval-AA 2.11844.00 / 3000自报
Global-MMLU92.1%自报

Healthcare

HealthBench Professional (raw)77.1%自报
HealthBench (raw)69.4%自报
HealthBench Professional69.2%自报
HealthBench65.4%自报
PhysicianBench63.2%自报

Language

MILU91.6%自报

Legal

Legal Agent Benchmark11.7%自报

Math

ArXivMath (with tools)95.2%自报
ArXivMath86.8%自报

Multimodal

OSWorld 2.1 (partial)80.1%自报
OSWorld 2.1 (strict)43.5%自报

Reasoning

SWE-bench Multilingual90.3%自报
SWE-Bench ProPrinceton NLP (2024)81.3%自报
Terminal-Bench 4.070.6%自报
Humanity's Last Exam (with tools)64.5%自报
FrontierCode 1.1 (Extended)64.4%自报
Terminal-Bench-Science 0.159.9%自报
Humanity's Last Exam (no tools, text-only)56.9%自报
SWE-Bench Multimodal54.3%自报
FrontierCode 1.152.1%自报

Science

LatchBio SingleCellBench59.1%自报

Vision

Chartography90.2%自报
Chartography (no tools)61.6%自报

AA 评测指数

(Artificial Analysis)
Lcr(Artificial Analysis)
82.7
Terminalbench V4 0
63.6
Scicode(UIUC + Argonne National Lab (2024))
61.0
Intelligence Index(Artificial Analysis)
56.0
Hle(Center for AI Safety + Scale AI (2025))
55.0

LLM Stats 分类评分

(LLM Stats (zeroeval))
Productivity
100
Agents
100
Reasoning
100
General
100
Language
90
Science
90
Biology
90
Multimodal
80
Vision
80
Math
70
Healthcare
70
Code
70
Tool Calling
60
Legal
10

定价

输入价格$2 / 1M tokens
输出价格$10 / 1M tokens
混合价格(3:1)$4 / 1M tokens
缓存读取价格$0.2 / 1M tokens
缓存写入价格$2.5 / 1M tokens

速度

Tokens/秒146.3
首Token延迟308.29s
首回答延迟308.29s

供应商价格排行

供应商价格排行

1 个供应商

供应商输入输出
1Anthropic
$0
$0.00001

比较该模型在不同 API 供应商之间的定价。

外部链接