Qwen2.5 Instruct 72B
AlibabaQwenオープンウエイトQwen · 商用利用可
説明
Qwen2.5-72B-Instruct is an instruction-tuned 72 billion parameter language model, part of the Qwen2.5 series. It is designed to follow instructions, generate long texts (over 8K tokens), understand structured data (e.g., tables), and generate structured outputs, especially JSON. The model supports multilingual capabilities across over 29 languages.
リリース日
2024-09-19
パラメータ
72.7B
コンテキスト長
131K
モダリティ
text
能力レーダー
26
general
28
coding
29
reasoning
35
science
29
agents
0
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| コーディングランキング | 513 | 19.0 | AA |
| 総合ランキング | 422 | 33.0 | AA |
| 科学 | 517 | 25.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Chat
MT-Bench
0.94 / 100自己申告
IFEvalGoogle Research (2023)
84.1%自己申告
General
AlignBench
81.6%自己申告
Arena Hard
81.2%自己申告
MultiPL-E
75.1%自己申告
Language
MMLU-Redux
86.8%自己申告
MMLU-Pro
71.1%自己申告
Math
GSM8k
95.8%自己申告
MATH
83.1%自己申告
LiveBench
52.3%自己申告
Reasoning
MBPP
0.88 / 100自己申告
HumanEvalOpenAI (2021)
86.6%自己申告
LiveCodeBench
55.5%自己申告
GPQANYU + Cohere + Anthropic (2023)
49.0%自己申告
AA評価指数
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))85.8
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))72.0
Gpqa(NYU + Cohere + Anthropic (2023))49.1
Ifbench(Google Research (2023))36.9
Tau2(Sierra + U Toronto + Vector Institute (2025))34.5
Livecodebench(UC Berkeley + MIT + Cornell (2024))27.6
Aime(MAA (Mathematical Association of America))16.0
Aime 25(MAA (Mathematical Association of America))14.0
Math Index(Artificial Analysis)14.0
Intelligence Index(Artificial Analysis)7.7
Terminalbench Hard(Stanford × Laude Institute (2026))4.5
Hle(Center for AI Safety + Scale AI (2025))3.6
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Chat90
Roleplay90
Communication90
Creativity90
Instruction Following80
Language80
Math80
Reasoning80
Structured Output80
General80
Writing80
Legal70
Finance70
Healthcare70
Code70
Physics50
Biology50
Chemistry50
価格設定
入力価格$0.475 / 1Mトークン
出力価格$0.495 / 1Mトークン
混合価格(3:1)$0.48 / 1Mトークン
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
3 プロバイダー
最安: DeepInfra最高: Alibaba (China)
プロバイダー入力出力
1DeepInfra最安
$0
$0
2Alibabaプライマリ
$0.475
$0.495
3Alibaba (China)
$0.574
$1.721
このモデルの異なるAPIプロバイダー間の価格を比較。