Claude Opus 4.6 (Non-reasoning, High Effort)
AnthropicClaudeProprietary
説明
Claude Opus 4.6 is Anthropic's most intelligent model, improving on its predecessor's coding skills with more careful planning, longer agentic task sustenance, more reliable operation in larger codebases, and better code review and debugging skills. First Opus-class model with 1M token context window (beta), 128K output tokens, and adaptive thinking. Features effort controls (low/medium/high/max) and context compaction for long-running tasks. State-of-the-art on Terminal-Bench 2.0, Humanity's Last Exam, GDPval-AA, and BrowseComp. Pricing: $5/$25 per million tokens (input/output).
リリース日
2026-02-05
パラメータ
—
コンテキスト長
1.0M
モダリティ
image, pdf, text
能力レーダー
35
general
46
coding
84
reasoning
58
science
80
agents
80
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| エージェント能力 | 23 | 57.0 | LS |
| コーディングランキング | 77 | 75.0 | AA |
| 総合ランキング | 130 | 64.0 | AA |
| マルチモーダルランキング | 43 | 47.0 | LS |
| 科学 | 113 | 66.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Agents
Vending-Bench 2
801759.0%自己申告
DeepSearchQA
91.3%自己申告
BrowseCompOpenAI (2025)
84.0%自己申告
CyberGym
73.8%自己申告
OSWorld
72.7%自己申告
Terminal-Bench 2.0Stanford × Laude Institute (2026)
65.4%自己申告
MCP Atlas
62.7%自己申告
Finance Agent
60.7%自己申告
FrontierSWE
56.0%
OpenRCA
34.9%自己申告
Legal Agent Benchmark
4.2%
Biology
GPQANYU + Cohere + Anthropic (2023)
91.3%自己申告
Code
SWE-Bench Verified
80.8%自己申告
SWE-bench Multilingual
77.8%自己申告
Communication
Tau2 Telecom
99.3%自己申告
Tau2 Retail
91.9%自己申告
General
MMMLU
91.1%自己申告
MMMU-Pro
77.3%自己申告
LiveBench
76.3%
MRCR v2 (8-needle)
76.0%自己申告
Healthcare
FigQA
78.3%自己申告
Long Context
Graphwalks parents >128k
95.4%自己申告
Graphwalks BFS >128k
61.5%自己申告
Math
AIME 2025
99.8%自己申告
Humanity's Last Exam
53.1%自己申告
Multimodal
CharXiv-R
77.4%自己申告
Reasoning
ARC-AGI v2
68.8%自己申告
AA評価指数
(Artificial Analysis)Intelligence Index(Artificial Analysis)38.8
Tau2(Sierra + U Toronto + Vector Institute (2025))0.8
Gpqa(NYU + Cohere + Anthropic (2023))0.8
Lcr(Artificial Analysis)0.6
Terminalbench Hard(Stanford × Laude Institute (2026))0.5
Scicode(UIUC + Argonne National Lab (2024))0.5
Ifbench(Google Research (2023))0.4
Hle(Center for AI Safety + Scale AI (2025))0.2
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Agents100
Reasoning100
General100
Communication100
Physics90
Search90
Language90
Biology90
Chemistry90
Long Context80
Math80
Multimodal80
Safety80
Spatial Reasoning80
Frontend Development80
Healthcare80
Tool Calling80
Code70
Vision70
Finance60
Legal0
価格設定
入力価格$5 / 1Mトークン
出力価格$25 / 1Mトークン
混合価格(3:1)$10 / 1Mトークン
キャッシュ読み取り価格$0.5 / 1Mトークン
キャッシュ書き込み価格$6.25 / 1Mトークン
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
16 プロバイダー
最安: Anthropic最高: Venice AI
プロバイダー入力出力
1Anthropic最安
$0.00001
$0.00003
2302.AI
$5
$25
3Abacus
$5
$25
4Jiekou.AI
$5
$25
5OpenCode Zen
$5
$25
6FrogBot
$5
$25
7AIHubMix
$5
$25
8Azure Cognitive Services
$5
$25
9Requesty
$5
$25
10LLM Gateway
$5
$25
11Azure
$5
$25
12Auriko
$5
$25
13FreeModel
$5
$25
14Neon
$5
$25
15Pioneer
$5
$25
16Venice AI
$6
$30
このモデルの異なるAPIプロバイダー間の価格を比較。