Перейти к основному содержанию

Claude Opus 4.8 (Adaptive Reasoning, Max Effort)

AnthropicClaudeProprietary

Описание

Claude Opus 4.8 is Anthropic's upgrade to Opus 4.7 and its most capable general-access model at release, with improvements across software engineering, agentic tool use, reasoning, computer use, and knowledge-work benchmarks while shipping at the same price ($5/$25 per million input/output tokens). Performance gains include SWE-Bench Verified (88.6%), SWE-Bench Pro (69.2%), Terminal-Bench 2.1 (74.6%), GPQA Diamond (93.6%), USAMO 2026 (96.7%), Humanity's Last Exam with tools (57.9%), OSWorld-Verified (83.4%), BrowseComp (84.3% single-agent, 88.5% multi-agent), MCP-Atlas (82.2%), and GDPval-AA (1890 Elo). The alignment assessment reports honesty improvements with around a four-fold drop in letting flaws in self-written code pass unremarked, a 17-fold drop relative to Sonnet 4.6 on dishonest agentic code summaries, and broadly improved adherence to Claude's constitution. The model defaults to high effort and exposes new 'extra' (xhigh) and 'max' levels for harder problems. Launches alongside Claude Code dynamic workflows (parallel subagents that plan, execute, and verify codebase-scale migrations), effort control in claude.ai and Cowork, and a Messages API extension that accepts system entries inside the messages array so harnesses can update instructions mid-task without breaking the prompt cache. Fast mode runs at 2.5× speed at $10/$50 per million input/output tokens, three times cheaper than fast mode on previous models. Available across Claude products, the Claude API as `claude-opus-4-8`, Amazon Bedrock, Google Cloud Vertex AI, and Microsoft Foundry.

Дата выхода
2026-05-28
Параметры
Длина контекста
1.0M
Модальности
image, pdf, text

Радар способностей

55
general
71
coding
92
reasoning
70
science
70
agents
80
multimodal

Рейтинги

Оценки бенчмарков (LLM Stats)

(LLM Stats (zeroeval))

Agents

OfficeQA Pro66.2%Сам.
Toolathlon59.9%Сам.
Finance Agent v253.9%
Finance Agent53.9%Сам.

Code

CyberGym78.8%Сам.
FrontierSWE75.0%
DeepSWE 1.159.0%

General

Include87.6%Сам.

Healthcare

HealthBench Professional55.8%Сам.

Math

LiveBench77.2%

Multimodal

OSWorld-Verified83.4%Сам.

Reasoning

GPQANYU + Cohere + Anthropic (2023)93.6%Сам.
CharXiv-R89.9%Сам.
SWE-Bench Verified88.6%Сам.
SWE-bench Multilingual84.4%Сам.
BrowseCompOpenAI (2025)84.3%Сам.
Graphwalks parents >128k83.3%Сам.
MCP Atlas82.2%Сам.
Terminal-Bench 2.0Stanford × Laude Institute (2026)74.6%Сам.
SWE-Bench ProPrinceton NLP (2024)69.2%Сам.
Graphwalks BFS >128k68.1%Сам.
Humanity's Last Exam57.9%Сам.
FrontierCode 1.146.5%
SWE-Bench Multimodal38.4%Сам.

Search

DeepSearchQA93.1%Сам.

Vision

ScreenSpot Pro87.9%Сам.

Индексы оценки AA

(Artificial Analysis)
Tau2(Sierra + U Toronto + Vector Institute (2025))
94.4
Gpqa(NYU + Cohere + Anthropic (2023))
92.0
Terminalbench V2 1
84.6
Coding Index(Artificial Analysis)
74.3
Lcr(Artificial Analysis)
73.0
Ifbench(Google Research (2023))
62.2
Terminalbench Hard(Stanford × Laude Institute (2026))
58.3
Intelligence Index(Artificial Analysis)
57.3
Scicode(UIUC + Argonne National Lab (2024))
53.5
Hle(Center for AI Safety + Scale AI (2025))
48.7
Tau Banking
34.2

Оценки категорий LLM Stats

(LLM Stats (zeroeval))
Physics
90
Search
90
Frontend Development
90
Grounding
90
Biology
90
Chemistry
90
Long Context
80
Safety
80
Spatial Reasoning
80
Math
70
Multimodal
70
Reasoning
70
General
70
Agents
70
Code
70
Tool Calling
70
Vision
70
Healthcare
60
Finance
50

Цены

Цена ввода$5 / 1M токенов
Цена вывода$25 / 1M токенов
Смешанная цена (3:1)$10 / 1M токенов
Цена чтения кэша$0.5 / 1M токенов
Цена записи кэша$6.25 / 1M токенов

Скорость

Токенов/сек0.0
Задержка первого токена0.00s
Время до первого ответа0.00s

Рейтинг цен провайдеров

Рейтинг цен провайдеров

30 провайдеров

Самый дешевый: AnthropicСамый дорогой: Venice AI
ПровайдерВводВывод
1AnthropicСамый дешевый
$0.00001
$0.00003
2UnoRouter
$0.425
$2.125
3Xpersona
$1.5
$9.25
4Poe
$4.2929
$21.4646
5NanoGPT
$5
$25
6Abacus
$5
$25
7OpenRouter
$5
$25
8ZenMux
$5
$25
9Kilo Gateway
$5
$25
10Cloudflare AI Gateway
$5
$25
11OpenCode Zen
$5
$25
12AIHubMix
$5
$25
13Azure Cognitive Services
$5
$25
14Vertex (Anthropic)
$5
$25
15Requesty
$5
$25
16Vercel AI Gateway
$5
$25
17DevPass (LLM Gateway)
$5
$25
18Vertex
$5
$25
19Azure
$5
$25
20FastRouter
$5
$25
21GMI Cloud
$5
$25
22OrcaRouter
$5
$25
23routing.run
$5
$25
24FreeModel
$5
$25
25Neon
$5
$25
26Pioneer
$5
$25
27DaoXE
$5
$25
28Ofox
$5
$25
29Modelis
$5
$25
30Venice AI
$6
$30

Сравнение цен разных API-провайдеров для этой модели.

Внешние ссылки