跳轉到主要內容

Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback)

AnthropicClaude

描述

Claude Fable 5 is Anthropic's generally available deployment of the same underlying weights as the restricted Claude Mythos 5 model. Anthropic positions the model as a new capability tier above Opus: Mythos 5 is the unsafeguarded trusted-access configuration, while Fable 5 adds production safeguards for public use. The system card reports that Mythos 5 is Anthropic's most capable model to date, with state-of-the-art results across software engineering, long-context agentic tasks, vision, knowledge work, healthcare, cyber, and life-sciences evaluations; Fable 5 matches the underlying model where safeguards do not trigger and behaves closer to Claude Opus 4.8 in safeguarded domains. Fable-specific results include SWE-Bench Verified 95.0%, SWE-Bench Pro 80.0%, Terminal-Bench 2.1 84.3% mean reward with 20.9% of trials hitting a safety fallback, OSWorld-Verified 85.0%, FrontierCode Diamond 29.3%, FrontierCode Main 46.3%, GDPval-AA 1932 Elo, GDP.pdf 29.8%, Blueprint-Bench 2 38.6%, and AutomationBench 17.4%. Fable 5 uses classifiers for cybersecurity, biology and chemistry, and distillation attempts; in Claude client apps these requests fall back to the latest Claude Opus model, while the Messages API blocks by default unless developers implement or opt into fallback. Separate competitive-use safeguards target frontier LLM development work and are estimated by Anthropic to affect about 0.03% of traffic. The model was trained on a proprietary mix of public web data, public and private datasets, and synthetic data. Anthropic prices Fable 5 at $10 per million input tokens and $50 per million output tokens, and exposes it through the Claude API as `claude-fable-5`.

發布日期
2026-06-09
參數規模
上下文長度
1.0M
支援模態
image, pdf, text

能力雷達圖

58
general
74
coding
93
reasoning
73
science估算
50
agents
60
multimodal

Science 在缺少專門科學評測時使用推理能力代理估算。

排行榜排名

領域#排名分數來源
程式碼能力榜6
96.0
AA
通用能力榜6
92.0
AA
科學能力1
99.0
AA

基準測試分數 (LLM Stats)

Agents

GDPval-AA1815.00 / 3000自報
FrontierSWE90.0%自報
OSWorld-Verified85.0%自報
Terminal-Bench 2.184.3%自報
SWE-Bench Pro80.0%自報
ExploitBench78.0%自報
DeepSWE 1.170.0%自報
Finance Agent v256.3%自報
FrontierCode 1.153.5%自報
FrontierCode46.3%自報
Blueprint-Bench 238.6%自報
AutomationBench17.4%自報
Legal Agent Benchmark13.3%自報

Biology

BioMysteryBench46.1%自報

Code

SWE-Bench Verified95.0%自報

General

LiveBench78.3%自報
GDP.pdf29.8%自報

Healthcare

HealthBench Professional66.0%自報

Math

Humanity's Last Exam64.5%自報

AA 評測指數

Coding Index
76.5
Intelligence Index
59.9
Tau2
1.0
Gpqa
0.9
Terminalbench V2 1
0.8
Lcr
0.7
Ifbench
0.6
Terminalbench Hard
0.6
Scicode
0.6
Hle
0.5
Tau Banking
0.3

LLM Stats 分類評分

Finance
100
Legal
100
Reasoning
100
Agents
100
General
100
Frontend Development
90
Safety
80
Math
70
Healthcare
70
Code
70
Multimodal
60
Vision
60
Tool Calling
50

定價

輸入價格$10 / 1M tokens
輸出價格$50 / 1M tokens
混合價格(3:1)$20 / 1M tokens
快取讀取價格$1 / 1M tokens
快取寫入價格$12.5 / 1M tokens

速度

Tokens/秒58.3
首Token延遲52.46s
首回答延遲52.46s

供應商價格排行

供應商價格排行

8 個供應商

最便宜: Anthropic最貴: Ofox
供應商輸入輸出
1Anthropic主要
$10
$50
2OpenRouter
$10
$50
3ZenMux
$10
$50
4Cloudflare AI Gateway
$10
$50
5Amazon Bedrock
$10
$50
6Vercel AI Gateway
$10
$50
7CrossModel
$10
$50
8Ofox
$10
$50

比較該模型在不同 API 供應商之間的定價。

外部連結