Claude 3 Opus
AnthropicClaudeProprietary
説明
Claude 3 Opus is Anthropic's most intelligent model, with best-in-market performance on highly complex tasks. It can navigate open-ended prompts and sight-unseen scenarios with remarkable fluency and human-like understanding, showing the outer limits of what's possible with generative AI.
リリース日
2024-03-04
パラメータ
—
コンテキスト長
—
モダリティ
image, text
能力レーダー
26
general
23
coding
31
reasoning
35
science
29
agents
80
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| コーディングランキング | 460 | 26.0 | AA |
| 総合ランキング | 435 | 32.0 | AA |
| 科学 | 530 | 24.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))General
MMLU
86.8%自己申告
Language
MMLU-Pro
68.5%自己申告
Math
GSM8k
95.0%自己申告
MGSM
90.7%自己申告
MATH
60.1%自己申告
Reasoning
ARC-C
96.4%自己申告
HellaSwagAI2 (2019)
95.4%自己申告
BIG-Bench Hard
86.8%自己申告
HumanEvalOpenAI (2021)
84.9%自己申告
DROP
83.1%自己申告
GPQANYU + Cohere + Anthropic (2023)
50.4%自己申告
AA評価指数
(Artificial Analysis)Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))69.6
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))64.1
Gpqa(NYU + Cohere + Anthropic (2023))48.9
Livecodebench(UC Berkeley + MIT + Cornell (2024))27.9
Coding Index(Artificial Analysis)19.5
Intelligence Index(Artificial Analysis)8.7
Aime(MAA (Mathematical Association of America))3.3
Hle(Center for AI Safety + Scale AI (2025))2.8
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Language80
Legal80
Math80
Reasoning80
Finance80
General80
Healthcare80
Code80
Physics50
Biology50
Chemistry50
価格設定
入力価格$15 / 1Mトークン
出力価格$75 / 1Mトークン
混合価格(3:1)$30 / 1Mトークン
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
1 プロバイダー
プロバイダー入力出力
1Anthropicプライマリ
$15
$75
このモデルの異なるAPIプロバイダー間の価格を比較。