Claude 3 Opus
AnthropicClaudeProprietary
विवरण
Claude 3 Opus is Anthropic's most intelligent model, with best-in-market performance on highly complex tasks. It can navigate open-ended prompts and sight-unseen scenarios with remarkable fluency and human-like understanding, showing the outer limits of what's possible with generative AI.
रिलीज़ तिथि
2024-03-04
पैरामीटर
—
संदर्भ लंबाई
—
मोडैलिटीज़
image, text
क्षमता रडार
26
general
23
coding
31
reasoning
35
science
29
agents
80
multimodal
रैंकिंग
| डोमेन | #रैंक | स्कोर | स्रोत |
|---|---|---|---|
| कोडिंग रैंकिंग | 460 | 26.0 | AA |
| सामान्य रैंकिंग | 435 | 32.0 | AA |
| विज्ञान | 530 | 24.0 | AA |
बेंचमार्क स्कोर (LLM Stats)
(LLM Stats (zeroeval))General
MMLU
86.8%स्वयं
Language
MMLU-Pro
68.5%स्वयं
Math
GSM8k
95.0%स्वयं
MGSM
90.7%स्वयं
MATH
60.1%स्वयं
Reasoning
ARC-C
96.4%स्वयं
HellaSwagAI2 (2019)
95.4%स्वयं
BIG-Bench Hard
86.8%स्वयं
HumanEvalOpenAI (2021)
84.9%स्वयं
DROP
83.1%स्वयं
GPQANYU + Cohere + Anthropic (2023)
50.4%स्वयं
AA मूल्यांकन सूचकांक
(Artificial Analysis)Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))69.6
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))64.1
Gpqa(NYU + Cohere + Anthropic (2023))48.9
Livecodebench(UC Berkeley + MIT + Cornell (2024))27.9
Coding Index(Artificial Analysis)19.5
Intelligence Index(Artificial Analysis)8.7
Aime(MAA (Mathematical Association of America))3.3
Hle(Center for AI Safety + Scale AI (2025))2.8
LLM Stats श्रेणी स्कोर
(LLM Stats (zeroeval))Language80
Legal80
Math80
Reasoning80
Finance80
General80
Healthcare80
Code80
Physics50
Biology50
Chemistry50
मूल्य निर्धारण
इनपुट मूल्य$15 / 1M टोकन
आउटपुट मूल्य$75 / 1M टोकन
मिश्रित मूल्य (3:1)$30 / 1M टोकन
गति
टोकन/सेकंड0.0
पहले टोकन में देरी0.00s
पहले उत्तर में देरी0.00s
प्रदाता मूल्य रैंकिंग
प्रदाता मूल्य रैंकिंग
1 प्रदाता
प्रदाताइनपुटआउटपुट
1Anthropicप्राथमिक
$15
$75
इस मॉडल के लिए विभिन्न API प्रदाताओं के मूल्य निर्धारण की तुलना करें।