메인 콘텐츠로 건너뛰기

Claude Opus 4.6 (Non-reasoning, High Effort)

AnthropicClaudeProprietary

설명

Claude Opus 4.6 is Anthropic's most intelligent model, improving on its predecessor's coding skills with more careful planning, longer agentic task sustenance, more reliable operation in larger codebases, and better code review and debugging skills. First Opus-class model with 1M token context window (beta), 128K output tokens, and adaptive thinking. Features effort controls (low/medium/high/max) and context compaction for long-running tasks. State-of-the-art on Terminal-Bench 2.0, Humanity's Last Exam, GDPval-AA, and BrowseComp. Pricing: $5/$25 per million tokens (input/output).

출시일
2026-02-05
파라미터
컨텍스트 길이
1.0M
모달리티
image, pdf, text

능력 레이더

35
general
46
coding
84
reasoning
58
science
80
agents
80
multimodal

랭킹

도메인#순위점수소스
에이전트형 역량25
57.0
LS
코딩 랭킹83
75.0
AA
종합 랭킹139
64.0
AA
멀티모달 랭킹46
46.0
LS
과학125
65.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Agents

Vending-Bench 2801759.0%자체 보고
DeepSearchQA91.3%자체 보고
BrowseCompOpenAI (2025)84.0%자체 보고
CyberGym73.8%자체 보고
OSWorld72.7%자체 보고
Terminal-Bench 2.0Stanford × Laude Institute (2026)65.4%자체 보고
MCP Atlas62.7%자체 보고
Finance Agent60.7%자체 보고
FrontierSWE56.0%
OpenRCA34.9%자체 보고
Legal Agent Benchmark4.2%

Biology

GPQANYU + Cohere + Anthropic (2023)91.3%자체 보고

Code

SWE-Bench Verified80.8%자체 보고
SWE-bench Multilingual77.8%자체 보고

Communication

Tau2 Telecom99.3%자체 보고
Tau2 Retail91.9%자체 보고

General

MMMLU91.1%자체 보고
MMMU-Pro77.3%자체 보고
LiveBench76.3%
MRCR v2 (8-needle)76.0%자체 보고

Healthcare

FigQA78.3%자체 보고

Long Context

Graphwalks parents >128k95.4%자체 보고
Graphwalks BFS >128k61.5%자체 보고

Math

AIME 202599.8%자체 보고
Humanity's Last Exam53.1%자체 보고

Multimodal

CharXiv-R77.4%자체 보고

Reasoning

ARC-AGI v268.8%자체 보고

AA 평가 지수

(Artificial Analysis)
Intelligence Index(Artificial Analysis)
38.8
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.8
Gpqa(NYU + Cohere + Anthropic (2023))
0.8
Lcr(Artificial Analysis)
0.6
Terminalbench Hard(Stanford × Laude Institute (2026))
0.5
Scicode(UIUC + Argonne National Lab (2024))
0.5
Ifbench(Google Research (2023))
0.4
Hle(Center for AI Safety + Scale AI (2025))
0.2

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Agents
100
Reasoning
100
General
100
Communication
100
Physics
90
Search
90
Language
90
Biology
90
Chemistry
90
Long Context
80
Math
80
Multimodal
80
Safety
80
Spatial Reasoning
80
Frontend Development
80
Healthcare
80
Tool Calling
80
Code
70
Vision
70
Finance
60
Legal
0

가격

입력 가격$5 / 1M 토큰
출력 가격$25 / 1M 토큰
혼합 가격 (3:1)$10 / 1M 토큰
캐시 읽기 가격$0.5 / 1M 토큰
캐시 쓰기 가격$6.25 / 1M 토큰

속도

토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s

공급자 가격 순위

공급자 가격 순위

1개 공급자

공급자입력출력
1Anthropic
$0.00001
$0.00003

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크