跳转到主要内容

Qwen2 Instruct 72B

AlibabaQwen开源权重tongyi-qianwen

描述

Qwen2-72B-Instruct is an instruction-tuned language model with 72 billion parameters, supporting a context length of up to 131,072 tokens. It's part of the new Qwen2 series, which has surpassed most open-source models and demonstrates competitiveness against proprietary models across various benchmarks.

发布日期
2024-06-07
参数规模
72.0B
上下文长度
支持模态

能力雷达图

22
general
17
coding
36
reasoning
25
science
30
agents
0
multimodal

排行榜排名

领域#排名分数来源
代码能力榜454
17.0
AA
通用能力榜453
28.0
AA
科学能力475
25.0
AA

基准测试分数 (LLM Stats)

(LLM Stats (zeroeval))

Biology

GPQANYU + Cohere + Anthropic (2023)42.4%自报

Code

HumanEvalOpenAI (2021)86.0%自报
EvalPlus0.79 / 100自报

Finance

MMLU82.3%自报
MMLU-Pro64.4%自报
TruthfulQA54.8%自报
TheoremQA44.4%自报

General

CMMLU90.1%自报
C-Eval83.8%自报
MBPP0.80 / 100自报
MultiPL-E69.2%自报
ARC-C68.9%自报

Language

Winogrande85.1%自报
BBH82.4%自报

Math

GSM8k91.1%自报
MATH59.7%自报

Reasoning

HellaSwagAI2 (2019)87.6%自报

AA 评测指数

(Artificial Analysis)
Intelligence Index(Artificial Analysis)
5.7
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
0.7
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.6
Gpqa(NYU + Cohere + Anthropic (2023))
0.4
Scicode(UIUC + Argonne National Lab (2024))
0.2
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.2
Aime(MAA (Mathematical Association of America))
0.1
Hle(Center for AI Safety + Scale AI (2025))
0.0

LLM Stats 分类评分

(LLM Stats (zeroeval))
Language
80
Code
80
Legal
70
Math
70
Reasoning
70
General
70
Healthcare
70
Finance
60
Physics
40
Biology
40
Chemistry
40

定价

输入价格免费
输出价格免费
混合价格(3:1)免费

速度

Tokens/秒0.0
首Token延迟0.00s
首回答延迟0.00s

供应商价格排行

暂无提供商数据

外部链接