메인 콘텐츠로 건너뛰기

Qwen2 Instruct 72B

AlibabaQwen오픈 웨이트tongyi-qianwen

설명

Qwen2-72B-Instruct is an instruction-tuned language model with 72 billion parameters, supporting a context length of up to 131,072 tokens. It's part of the new Qwen2 series, which has surpassed most open-source models and demonstrates competitiveness against proprietary models across various benchmarks.

출시일
2024-06-07
파라미터
72.0B
컨텍스트 길이
—
모달리티
—

능력 레이더

23
general
16
coding
36
reasoning
27
science
30
agents
0
multimodal

랭킹

도메인#순위점수소스
코딩 랭킹535
17.0
AA
종합 랭킹527
26.0
AA
과학584
18.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

General

C-Eval83.8%자체 보고
MMLU82.3%자체 보고
MultiPL-E69.2%자체 보고
TruthfulQA54.8%자체 보고

Language

CMMLU90.1%자체 보고
MMLU-Pro64.4%자체 보고

Math

GSM8k91.1%자체 보고
MATH59.7%자체 보고
TheoremQA44.4%자체 보고

Reasoning

HellaSwagAI2 (2019)87.6%자체 보고
HumanEvalOpenAI (2021)86.0%자체 보고
Winogrande85.1%자체 보고
BBH82.4%자체 보고
MBPP0.80 / 100자체 보고
EvalPlus0.79 / 100자체 보고
ARC-C68.9%자체 보고
GPQANYU + Cohere + Anthropic (2023)42.4%자체 보고

AA 평가 지수

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
70.1
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
62.2
Gpqa(NYU + Cohere + Anthropic (2023))
37.1
Livecodebench(UC Berkeley + MIT + Cornell (2024))
15.9
Aime(MAA (Mathematical Association of America))
14.7
Intelligence Index(Artificial Analysis)
6.3
Hle(Center for AI Safety + Scale AI (2025))
3.7

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Language
80
Code
80
Legal
70
Math
70
Reasoning
70
General
70
Healthcare
70
Finance
60
Physics
40
Biology
40
Chemistry
40

가격

입력 가격무료
출력 가격무료
혼합 가격 (3:1)무료

속도

토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s

공급자 가격 순위

프로바이더 데이터가 없습니다

외부 링크