Claude Fable 5.1
説明
Claude Fable 5.1 is Anthropic's generally available, production-safeguarded deployment of the same underlying weights as Claude Mythos 5.1 (trusted-access only via CVP/LSVP; not cataloged here). It targets demanding reasoning and long-horizon agentic work with text and image input, text output, multilingual and vision support, tool use, adaptive thinking always on (API/Claude Code default effort high; Cowork/claude.ai default medium), a 1M-token context window, and 128K max output on the sync Messages API. Reliable knowledge and training-data cutoffs are June 2026; comparative latency is slower than the rest of the current lineup. First-party pricing matches Fable 5 at $10/$50 per million input/output tokens, with cache reads cut to $0.25 per million tokens (0.025x base input, down from $1 / 0.1x on Fable 5); Anthropic estimates ~25% lower cost on typical workloads and ~45% on highly agentic ones. Self-reported launch benchmarks with production safeguards enabled include Terminal-Bench-Science 0.1 52.6% (SE ±3.5–4.5 pts), Terminal-Bench 4.0 55.8%, GDPval-AA v2 1853 Elo, OSWorld 2.0 77.9% partial / 41.7% strict (August 2026 task release), Humanity's Last Exam 60.9% no tools / 65.0% with tools, AutomationBench 31.4%, and CursorBench 3.2.0 73.4%. Available on the Claude API as `claude-fable-5-1`.
能力レーダー
専用の科学ベンチマークがない場合、Science は LLM Stats の科学スコアまたは推論能力から推定します。
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| エージェント能力 | 22 | 55.0 | LS |
| 数学的推論 | 1 | 97.0 | LB |
| マルチモーダルランキング | 11 | 65.0 | LS |
| 推論 | 1 | 92.0 | LB |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Agents
Code
General
Multimodal
Reasoning
AA評価指数
(Artificial Analysis)AA評価データがありません
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))価格設定
速度
速度データがありません
プロバイダー価格ランキング
プロバイダー価格ランキング
12 プロバイダー
このモデルの異なるAPIプロバイダー間の価格を比較。