Claude Opus 4.6 (Non-reasoning, High Effort)
AnthropicClaudeProprietary
विवरण
Claude Opus 4.6 is Anthropic's most intelligent model, improving on its predecessor's coding skills with more careful planning, longer agentic task sustenance, more reliable operation in larger codebases, and better code review and debugging skills. First Opus-class model with 1M token context window (beta), 128K output tokens, and adaptive thinking. Features effort controls (low/medium/high/max) and context compaction for long-running tasks. State-of-the-art on Terminal-Bench 2.0, Humanity's Last Exam, GDPval-AA, and BrowseComp. Pricing: $5/$25 per million tokens (input/output).
रिलीज़ तिथि
2026-02-05
पैरामीटर
—
संदर्भ लंबाई
1.0M
मोडैलिटीज़
image, pdf, text
क्षमता रडार
35
general
46
coding
84
reasoning
58
science
80
agents
80
multimodal
रैंकिंग
| डोमेन | #रैंक | स्कोर | स्रोत |
|---|---|---|---|
| एजेंटिक क्षमता | 25 | 57.0 | LS |
| कोडिंग रैंकिंग | 83 | 75.0 | AA |
| सामान्य रैंकिंग | 139 | 64.0 | AA |
| मल्टीमॉडल रैंकिंग | 46 | 46.0 | LS |
| विज्ञान | 125 | 65.0 | AA |
बेंचमार्क स्कोर (LLM Stats)
(LLM Stats (zeroeval))Agents
Vending-Bench 2
801759.0%स्वयं
DeepSearchQA
91.3%स्वयं
BrowseCompOpenAI (2025)
84.0%स्वयं
CyberGym
73.8%स्वयं
OSWorld
72.7%स्वयं
Terminal-Bench 2.0Stanford × Laude Institute (2026)
65.4%स्वयं
MCP Atlas
62.7%स्वयं
Finance Agent
60.7%स्वयं
FrontierSWE
56.0%
OpenRCA
34.9%स्वयं
Legal Agent Benchmark
4.2%
Biology
GPQANYU + Cohere + Anthropic (2023)
91.3%स्वयं
Code
SWE-Bench Verified
80.8%स्वयं
SWE-bench Multilingual
77.8%स्वयं
Communication
Tau2 Telecom
99.3%स्वयं
Tau2 Retail
91.9%स्वयं
General
MMMLU
91.1%स्वयं
MMMU-Pro
77.3%स्वयं
LiveBench
76.3%
MRCR v2 (8-needle)
76.0%स्वयं
Healthcare
FigQA
78.3%स्वयं
Long Context
Graphwalks parents >128k
95.4%स्वयं
Graphwalks BFS >128k
61.5%स्वयं
Math
AIME 2025
99.8%स्वयं
Humanity's Last Exam
53.1%स्वयं
Multimodal
CharXiv-R
77.4%स्वयं
Reasoning
ARC-AGI v2
68.8%स्वयं
AA मूल्यांकन सूचकांक
(Artificial Analysis)Intelligence Index(Artificial Analysis)38.8
Tau2(Sierra + U Toronto + Vector Institute (2025))0.8
Gpqa(NYU + Cohere + Anthropic (2023))0.8
Lcr(Artificial Analysis)0.6
Terminalbench Hard(Stanford × Laude Institute (2026))0.5
Scicode(UIUC + Argonne National Lab (2024))0.5
Ifbench(Google Research (2023))0.4
Hle(Center for AI Safety + Scale AI (2025))0.2
LLM Stats श्रेणी स्कोर
(LLM Stats (zeroeval))Agents100
Reasoning100
General100
Communication100
Physics90
Search90
Language90
Biology90
Chemistry90
Long Context80
Math80
Multimodal80
Safety80
Spatial Reasoning80
Frontend Development80
Healthcare80
Tool Calling80
Code70
Vision70
Finance60
Legal0
मूल्य निर्धारण
इनपुट मूल्य$5 / 1M टोकन
आउटपुट मूल्य$25 / 1M टोकन
मिश्रित मूल्य (3:1)$10 / 1M टोकन
कैश पठन मूल्य$0.5 / 1M टोकन
कैश लेखन मूल्य$6.25 / 1M टोकन
गति
टोकन/सेकंड0.0
पहले टोकन में देरी0.00s
पहले उत्तर में देरी0.00s
प्रदाता मूल्य रैंकिंग
प्रदाता मूल्य रैंकिंग
1 प्रदाता
प्रदाताइनपुटआउटपुट
1Anthropic
$0.00001
$0.00003
इस मॉडल के लिए विभिन्न API प्रदाताओं के मूल्य निर्धारण की तुलना करें।