跳轉到主要內容

Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback)

AnthropicClaudeProprietary

描述

Claude Opus 5.5 is the first model in Anthropic's Claude 5.5 family. It performs at Claude Fable 5.1 level on most work at about 40% lower cost than Claude Opus 5, and targets long-running agentic coding and knowledge work. It accepts text and image input with text output, supports multilingual and vision workloads, tool use, and adaptive thinking that is always on (API default effort medium). Context window is 1M tokens with 128K max output on the sync Messages API; reliable knowledge and training-data cutoffs are June 2026. First-party pricing is $4/$20 per million input/output tokens, with cache reads at $0.20 and 5-minute cache writes at $5 per million tokens. Fast mode is available in Claude Code and on the Claude Platform at $8/$40 per million input/output tokens with up to 2.5x speed (same model id; not a separate deployment). Capability scores are self-reported from the Claude Opus 5.5 System Card (Table 8.1.A and §8) and the launch announcement. Available on the Claude API as `claude-opus-5-5`.

發布日期
2026-09-22
參數規模
—
上下文長度
1.0M
支援模態
image, pdf, text

能力雷達圖

58
general
67
coding
70
reasoning
65
science
60
agents
80
multimodal

排行榜排名

領域#排名分數來源
智慧體能力模型榜5
69.0
LS
程式碼能力榜9
95.0
AA
通用能力榜1
100.0
AA
多模態榜13
65.0
LS
科學能力1
100.0
AA

基準測試分數 (LLM Stats)

(LLM Stats (zeroeval))

Agents

AA-Briefcase v1.11822.00 / 3000自報
Program Bench91.2%自報
OfficeQA78.9%自報
Toolathlon-Verified77.8%自報
OfficeQA Pro67.7%自報
AutomationBench40.0%自報

Biology

BioMysteryBench89.3%自報
LatchBio SpatialBench Verified72.0%自報

Code

BenchCAD (with Python tool)96.2%自報
DeepSWE 1.174.2%自報
BenchCAD73.0%自報
FrontierSWE V262.3%自報
CursorBench 4.057.8%自報

General

GDPval-AA 2.11846.00 / 3000自報
Global-MMLU94.3%自報

Healthcare

HealthBench Professional65.6%自報
HealthBench60.6%自報

Language

MILU93.1%自報

Legal

Legal Agent Benchmark8.3%自報

Math

ArXivMath91.2%自報

Multimodal

OSWorld 2.081.8%自報

Reasoning

SWE-bench Multilingual93.9%自報
SWE-Bench ProPrinceton NLP (2024)89.9%自報
Humanity's Last Exam (with tools, text-only)67.7%自報
Terminal-Bench 4.066.4%自報
Humanity's Last Exam (no tools, text-only)64.4%自報
SWE-Bench Multimodal61.4%自報
Terminal-Bench-Science 0.158.7%自報
FrontierCode 1.154.4%自報

Science

LatchBio SingleCellBench61.2%自報

Vision

Chartography89.0%自報

AA 評測指數

(Artificial Analysis)
Lcr(Artificial Analysis)
84.7
Scicode(UIUC + Argonne National Lab (2024))
66.9
Hle(Center for AI Safety + Scale AI (2025))
61.4
Intelligence Index(Artificial Analysis)
57.6

LLM Stats 分類評分

(LLM Stats (zeroeval))
Language
90
Science
90
Biology
90
Multimodal
80
Code
80
Vision
80
Math
70
Reasoning
70
General
70
Healthcare
60
Agents
60
Tool Calling
60
Legal
10

定價

輸入價格$4 / 1M tokens
輸出價格$20 / 1M tokens
混合價格(3:1)$8 / 1M tokens
快取讀取價格$0.2 / 1M tokens
快取寫入價格$5 / 1M tokens

速度

Tokens/秒95.5
首Token延遲388.27s
首回答延遲388.27s

供應商價格排行

供應商價格排行

1 個供應商

供應商輸入輸出
1Anthropic
$0
$0.00002

比較該模型在不同 API 供應商之間的定價。

外部連結