跳轉到主要內容

Claude Opus 4.8 (Adaptive Reasoning, Max Effort)

AnthropicClaudeProprietary

描述

Claude Opus 4.8 is Anthropic's upgrade to Opus 4.7 and its most capable general-access model at release, with improvements across software engineering, agentic tool use, reasoning, computer use, and knowledge-work benchmarks while shipping at the same price ($5/$25 per million input/output tokens). Performance gains include SWE-Bench Verified (88.6%), SWE-Bench Pro (69.2%), Terminal-Bench 2.1 (74.6%), GPQA Diamond (93.6%), USAMO 2026 (96.7%), Humanity's Last Exam with tools (57.9%), OSWorld-Verified (83.4%), BrowseComp (84.3% single-agent, 88.5% multi-agent), MCP-Atlas (82.2%), and GDPval-AA (1890 Elo). The alignment assessment reports honesty improvements with around a four-fold drop in letting flaws in self-written code pass unremarked, a 17-fold drop relative to Sonnet 4.6 on dishonest agentic code summaries, and broadly improved adherence to Claude's constitution. The model defaults to high effort and exposes new 'extra' (xhigh) and 'max' levels for harder problems. Launches alongside Claude Code dynamic workflows (parallel subagents that plan, execute, and verify codebase-scale migrations), effort control in claude.ai and Cowork, and a Messages API extension that accepts system entries inside the messages array so harnesses can update instructions mid-task without breaking the prompt cache. Fast mode runs at 2.5× speed at $10/$50 per million input/output tokens, three times cheaper than fast mode on previous models. Available across Claude products, the Claude API as `claude-opus-4-8`, Amazon Bedrock, Google Cloud Vertex AI, and Microsoft Foundry.

發布日期
2026-05-28
參數規模
上下文長度
1.0M
支援模態
image, pdf, text

能力雷達圖

55
general
71
coding
92
reasoning
70
science
70
agents
80
multimodal

排行榜排名

領域#排名分數來源
智慧體能力模型榜22
55.0
LS
程式碼能力榜27
92.0
AA
通用能力榜27
87.0
AA
數學推理6
94.0
LB
多模態榜8
69.0
LS
推理能力9
89.0
LB
科學能力15
91.0
AA

基準測試分數 (LLM Stats)

(LLM Stats (zeroeval))

Agents

OfficeQA Pro66.2%自報
Toolathlon59.9%自報
Finance Agent v253.9%
Finance Agent53.9%自報

Code

CyberGym78.8%自報
FrontierSWE75.0%
DeepSWE 1.159.0%

General

Include87.6%自報

Healthcare

HealthBench Professional55.8%自報

Math

LiveBench77.2%

Multimodal

OSWorld-Verified83.4%自報

Reasoning

GPQANYU + Cohere + Anthropic (2023)93.6%自報
CharXiv-R89.9%自報
SWE-Bench Verified88.6%自報
SWE-bench Multilingual84.4%自報
BrowseCompOpenAI (2025)84.3%自報
Graphwalks parents >128k83.3%自報
MCP Atlas82.2%自報
Terminal-Bench 2.0Stanford × Laude Institute (2026)74.6%自報
SWE-Bench ProPrinceton NLP (2024)69.2%自報
Graphwalks BFS >128k68.1%自報
Humanity's Last Exam57.9%自報
FrontierCode 1.146.5%
SWE-Bench Multimodal38.4%自報

Search

DeepSearchQA93.1%自報

Vision

ScreenSpot Pro87.9%自報

AA 評測指數

(Artificial Analysis)
Tau2(Sierra + U Toronto + Vector Institute (2025))
94.4
Gpqa(NYU + Cohere + Anthropic (2023))
92.0
Terminalbench V2 1
84.6
Coding Index(Artificial Analysis)
74.3
Lcr(Artificial Analysis)
73.0
Ifbench(Google Research (2023))
62.2
Terminalbench Hard(Stanford × Laude Institute (2026))
58.3
Intelligence Index(Artificial Analysis)
57.3
Scicode(UIUC + Argonne National Lab (2024))
53.5
Hle(Center for AI Safety + Scale AI (2025))
48.7
Tau Banking
34.2

LLM Stats 分類評分

(LLM Stats (zeroeval))
Physics
90
Search
90
Frontend Development
90
Grounding
90
Biology
90
Chemistry
90
Long Context
80
Safety
80
Spatial Reasoning
80
Math
70
Multimodal
70
Reasoning
70
General
70
Agents
70
Code
70
Tool Calling
70
Vision
70
Healthcare
60
Finance
50

定價

輸入價格$5 / 1M tokens
輸出價格$25 / 1M tokens
混合價格(3:1)$10 / 1M tokens
快取讀取價格$0.5 / 1M tokens
快取寫入價格$6.25 / 1M tokens

速度

Tokens/秒0.0
首Token延遲0.00s
首回答延遲0.00s

供應商價格排行

供應商價格排行

30 個供應商

最便宜: Anthropic最貴: Venice AI
供應商輸入輸出
1Anthropic最便宜
$0.00001
$0.00003
2UnoRouter
$0.425
$2.125
3Xpersona
$1.5
$9.25
4Poe
$4.2929
$21.4646
5NanoGPT
$5
$25
6Abacus
$5
$25
7OpenRouter
$5
$25
8ZenMux
$5
$25
9Kilo Gateway
$5
$25
10Cloudflare AI Gateway
$5
$25
11OpenCode Zen
$5
$25
12AIHubMix
$5
$25
13Azure Cognitive Services
$5
$25
14Vertex (Anthropic)
$5
$25
15Requesty
$5
$25
16Vercel AI Gateway
$5
$25
17DevPass (LLM Gateway)
$5
$25
18Vertex
$5
$25
19Azure
$5
$25
20FastRouter
$5
$25
21GMI Cloud
$5
$25
22OrcaRouter
$5
$25
23routing.run
$5
$25
24FreeModel
$5
$25
25Neon
$5
$25
26Pioneer
$5
$25
27DaoXE
$5
$25
28Ofox
$5
$25
29Modelis
$5
$25
30Venice AI
$6
$30

比較該模型在不同 API 供應商之間的定價。

外部連結