跳转到主要内容

Claude 3.5 Haiku

AnthropicClaudeProprietary

描述

Claude 3.5 Haiku is Anthropic's fastest model, delivering advanced coding, tool use, and reasoning capabilities at an accessible price. It excels at user-facing products, specialized sub-agent tasks, and generating personalized experiences from large data volumes. The model is particularly well-suited for code completions, interactive chatbots, data extraction, and real-time content moderation.

发布日期
2024-10-22
参数规模
—
上下文长度
—
支持模态
text

能力雷达图

24
general
22
coding
31
reasoning
29
science
40
agents
0
multimodal

排行榜排名

领域#排名分数来源
代码能力榜498
21.0
AA
通用能力榜432
32.0
AA
科学能力571
20.0
AA

基准测试分数 (LLM Stats)

(LLM Stats (zeroeval))

Chat

TAU-bench Retail51.0%自报

Language

MMLU-Pro65.0%自报

Math

MGSM85.6%自报
MATH69.4%自报

Reasoning

HumanEvalOpenAI (2021)88.1%自报
DROP83.1%自报
GPQANYU + Cohere + Anthropic (2023)41.6%自报
SWE-Bench Verified40.6%自报
TAU-bench Airline22.8%自报

AA 评测指数

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
72.1
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
63.4
Ifbench(Google Research (2023))
42.8
Gpqa(NYU + Cohere + Anthropic (2023))
40.8
Livecodebench(UC Berkeley + MIT + Cornell (2024))
31.4
Lcr(Artificial Analysis)
27.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
24.6
Coding Index(Artificial Analysis)
15.9
Terminalbench V2 1
10.1
Intelligence Index(Artificial Analysis)
8.9
Hle(Center for AI Safety + Scale AI (2025))
3.6
Aime(MAA (Mathematical Association of America))
3.3
Terminalbench Hard(Stanford × Laude Institute (2026))
2.3

LLM Stats 分类评分

(LLM Stats (zeroeval))
Math
80
Language
70
Legal
70
Finance
70
Healthcare
70
Reasoning
60
General
60
Code
60
Chat
50
Physics
40
Frontend Development
40
Biology
40
Chemistry
40
Communication
40
Tool Calling
40

定价

输入价格免费
输出价格免费
混合价格(3:1)免费

速度

Tokens/秒0.0
首Token延迟0.00s
首回答延迟0.00s

供应商价格排行

供应商价格排行

1 个供应商

供应商输入输出
1OpenCode Zen
$0.8
$4

比较该模型在不同 API 供应商之间的定价。

外部链接