Qwen3.5 397B A17B (Non-reasoning)
AlibabaQwenオープンウエイトApache 2.0 · 商用利用可
説明
Qwen3.5-397B-A17B is Qwen's flagship Mixture-of-Experts model with 397 billion total parameters and 17 billion activated parameters. It delivers state-of-the-art performance across knowledge, reasoning, coding, mathematics, multilingual understanding, instruction following, long context, and agent tasks.
リリース日
2026-02-16
パラメータ
397.0B
コンテキスト長
262K
モダリティ
audio, image, text, video
能力レーダー
21
general
70
coding
86
reasoning
66
science
60
agents
70
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| エージェント能力 | 17 | 62.0 | LS |
| コーディングランキング | 216 | 64.0 | AA |
| 総合ランキング | 211 | 52.0 | AA |
| 科学 | 195 | 60.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
86.7%自己申告
VITA-Bench
49.7%自己申告
MCP-Mark
46.1%自己申告
Toolathlon
38.3%自己申告
DeepPlanning
34.3%自己申告
Chat
IFEvalGoogle Research (2023)
92.6%自己申告
Multi-Challenge
67.6%自己申告
Code
SecCodeBench
68.3%自己申告
General
C-Eval
93.0%自己申告
MAXIFE
88.2%自己申告
Include
85.6%自己申告
NOVA-63
59.1%自己申告
Instruction Following
IFBench
76.5%自己申告
Language
MMLU-Redux
94.9%自己申告
MMMLU
88.5%自己申告
MMLU-Pro
87.8%自己申告
MMLU-ProX
84.7%自己申告
WMT24++
78.9%自己申告
Long Context
LongBench v2
63.2%自己申告
Math
HMMT 2025
94.8%自己申告
HMMT25
92.7%自己申告
AIME 2026
91.3%自己申告
IMO-AnswerBench
80.9%自己申告
PolyMATH
73.3%自己申告
Reasoning
Global PIQA
89.8%自己申告
GPQANYU + Cohere + Anthropic (2023)
88.4%自己申告
LiveCodeBench v6
83.6%自己申告
SWE-Bench Verified
76.4%自己申告
SuperGPQA
70.4%自己申告
BrowseComp-zh
70.3%自己申告
SWE-bench Multilingual
69.3%自己申告
BrowseCompOpenAI (2025)
69.0%自己申告
AA-LCR
68.7%自己申告
Terminal-Bench 2.0Stanford × Laude Institute (2026)
52.5%自己申告
Seal-0
46.9%自己申告
Humanity's Last Exam
28.7%自己申告
Search
WideSearch
74.0%自己申告
Tool Calling
BFCL-V4
72.9%自己申告
AA評価指数
(Artificial Analysis)Gpqa(NYU + Cohere + Anthropic (2023))86.1
Tau2(Sierra + U Toronto + Vector Institute (2025))83.9
Lcr(Artificial Analysis)64.3
Ifbench(Google Research (2023))51.6
Terminalbench Hard(Stanford × Laude Institute (2026))35.6
Intelligence Index(Artificial Analysis)21.4
Hle(Center for AI Safety + Scale AI (2025))19.8
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Language90
Biology90
Chat80
Instruction Following80
Legal80
Math80
Physics80
Structured Output80
Finance80
Frontend Development80
Healthcare80
Chemistry80
Long Context70
Multimodal70
Reasoning70
Search70
Spatial Reasoning70
General70
Code70
Communication70
Economics70
Agents60
Tool Calling60
Vision50
価格設定
入力価格$0.6 / 1Mトークン
出力価格$3.6 / 1Mトークン
混合価格(3:1)$1.35 / 1Mトークン
速度
トークン/秒86.1
初トークン遅延1.70s
初回答遅延1.70s
プロバイダー価格ランキング
プロバイダー価格ランキング
2 プロバイダー
最安: DeepInfra最高: Alibaba
プロバイダー入力出力
1DeepInfra最安
$0
$0
2Alibabaプライマリ
$0.6
$3.6
このモデルの異なるAPIプロバイダー間の価格を比較。