メインコンテンツへスキップ

Claude Opus 4.8 (Adaptive Reasoning, Max Effort)

AnthropicClaudeProprietary

説明

Claude Opus 4.8 is Anthropic's upgrade to Opus 4.7 and its most capable general-access model at release, with improvements across software engineering, agentic tool use, reasoning, computer use, and knowledge-work benchmarks while shipping at the same price ($5/$25 per million input/output tokens). Performance gains include SWE-Bench Verified (88.6%), SWE-Bench Pro (69.2%), Terminal-Bench 2.1 (74.6%), GPQA Diamond (93.6%), USAMO 2026 (96.7%), Humanity's Last Exam with tools (57.9%), OSWorld-Verified (83.4%), BrowseComp (84.3% single-agent, 88.5% multi-agent), MCP-Atlas (82.2%), and GDPval-AA (1890 Elo). The alignment assessment reports honesty improvements with around a four-fold drop in letting flaws in self-written code pass unremarked, a 17-fold drop relative to Sonnet 4.6 on dishonest agentic code summaries, and broadly improved adherence to Claude's constitution. The model defaults to high effort and exposes new 'extra' (xhigh) and 'max' levels for harder problems. Launches alongside Claude Code dynamic workflows (parallel subagents that plan, execute, and verify codebase-scale migrations), effort control in claude.ai and Cowork, and a Messages API extension that accepts system entries inside the messages array so harnesses can update instructions mid-task without breaking the prompt cache. Fast mode runs at 2.5× speed at $10/$50 per million input/output tokens, three times cheaper than fast mode on previous models. Available across Claude products, the Claude API as `claude-opus-4-8`, Amazon Bedrock, Google Cloud Vertex AI, and Microsoft Foundry.

リリース日
2026-05-28
パラメータ
コンテキスト長
1.0M
モダリティ
image, pdf, text

能力レーダー

55
general
71
coding
92
reasoning
70
science
70
agents
80
multimodal

ランキング

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Agents

OfficeQA Pro66.2%自己申告
Toolathlon59.9%自己申告
Finance Agent v253.9%
Finance Agent53.9%自己申告

Code

CyberGym78.8%自己申告
FrontierSWE75.0%
DeepSWE 1.159.0%

General

Include87.6%自己申告

Healthcare

HealthBench Professional55.8%自己申告

Math

LiveBench77.2%

Multimodal

OSWorld-Verified83.4%自己申告

Reasoning

GPQANYU + Cohere + Anthropic (2023)93.6%自己申告
CharXiv-R89.9%自己申告
SWE-Bench Verified88.6%自己申告
SWE-bench Multilingual84.4%自己申告
BrowseCompOpenAI (2025)84.3%自己申告
Graphwalks parents >128k83.3%自己申告
MCP Atlas82.2%自己申告
Terminal-Bench 2.0Stanford × Laude Institute (2026)74.6%自己申告
SWE-Bench ProPrinceton NLP (2024)69.2%自己申告
Graphwalks BFS >128k68.1%自己申告
Humanity's Last Exam57.9%自己申告
FrontierCode 1.146.5%
SWE-Bench Multimodal38.4%自己申告

Search

DeepSearchQA93.1%自己申告

Vision

ScreenSpot Pro87.9%自己申告

AA評価指数

(Artificial Analysis)
Tau2(Sierra + U Toronto + Vector Institute (2025))
94.4
Gpqa(NYU + Cohere + Anthropic (2023))
92.0
Terminalbench V2 1
84.6
Coding Index(Artificial Analysis)
74.3
Lcr(Artificial Analysis)
73.0
Ifbench(Google Research (2023))
62.2
Terminalbench Hard(Stanford × Laude Institute (2026))
58.3
Intelligence Index(Artificial Analysis)
57.3
Scicode(UIUC + Argonne National Lab (2024))
53.5
Hle(Center for AI Safety + Scale AI (2025))
48.7
Tau Banking
34.2

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Physics
90
Search
90
Frontend Development
90
Grounding
90
Biology
90
Chemistry
90
Long Context
80
Safety
80
Spatial Reasoning
80
Math
70
Multimodal
70
Reasoning
70
General
70
Agents
70
Code
70
Tool Calling
70
Vision
70
Healthcare
60
Finance
50

価格設定

入力価格$5 / 1Mトークン
出力価格$25 / 1Mトークン
混合価格(3:1)$10 / 1Mトークン
キャッシュ読み取り価格$0.5 / 1Mトークン
キャッシュ書き込み価格$6.25 / 1Mトークン

速度

トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s

プロバイダー価格ランキング

プロバイダー価格ランキング

30 プロバイダー

最安: Anthropic最高: Venice AI
プロバイダー入力出力
1Anthropic最安
$0.00001
$0.00003
2UnoRouter
$0.425
$2.125
3Xpersona
$1.5
$9.25
4Poe
$4.2929
$21.4646
5NanoGPT
$5
$25
6Abacus
$5
$25
7OpenRouter
$5
$25
8ZenMux
$5
$25
9Kilo Gateway
$5
$25
10Cloudflare AI Gateway
$5
$25
11OpenCode Zen
$5
$25
12AIHubMix
$5
$25
13Azure Cognitive Services
$5
$25
14Vertex (Anthropic)
$5
$25
15Requesty
$5
$25
16Vercel AI Gateway
$5
$25
17DevPass (LLM Gateway)
$5
$25
18Vertex
$5
$25
19Azure
$5
$25
20FastRouter
$5
$25
21GMI Cloud
$5
$25
22OrcaRouter
$5
$25
23routing.run
$5
$25
24FreeModel
$5
$25
25Neon
$5
$25
26Pioneer
$5
$25
27DaoXE
$5
$25
28Ofox
$5
$25
29Modelis
$5
$25
30Venice AI
$6
$30

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク