Qwen3.5 2B (Reasoning)
AlibabaQwenオープンウエイトApache 2.0 · 商用利用可
説明
Qwen3.5-2B is a 2 billion parameter vision-language model using Gated DeltaNet hybrid architecture with a 3:1 ratio of linear attention to full softmax attention. It supports 262K native context length and features both thinking and non-thinking modes.
リリース日
2026-03-02
パラメータ
2.0B
コンテキスト長
—
モダリティ
—
能力レーダー
6
general
3
coding
46
reasoning
32
science
50
agents
30
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| エージェント能力 | 106 | 31.0 | LS |
| コーディングランキング | 584 | 10.0 | AA |
| 総合ランキング | 501 | 28.0 | AA |
| 科学 | 551 | 22.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
48.8%自己申告
Chat
IFEvalGoogle Research (2023)
78.6%自己申告
Multi-Challenge
33.7%自己申告
General
C-Eval
73.2%自己申告
MAXIFE
60.6%自己申告
Include
55.4%自己申告
NOVA-63
46.4%自己申告
Instruction Following
IFBench
41.3%自己申告
Language
MMLU-Redux
79.6%自己申告
MMLU-Pro
66.5%自己申告
MMMLU
63.1%自己申告
MMLU-ProX
52.3%自己申告
WMT24++
45.8%自己申告
Long Context
LongBench v2
38.7%自己申告
Math
PolyMATH
26.1%自己申告
Reasoning
Global PIQA
69.3%自己申告
GPQANYU + Cohere + Anthropic (2023)
51.6%自己申告
SuperGPQA
37.5%自己申告
AA-LCR
25.6%自己申告
Tool Calling
BFCL-V4
43.6%自己申告
AA評価指数
(Artificial Analysis)Tau2(Sierra + U Toronto + Vector Institute (2025))69.0
Gpqa(NYU + Cohere + Anthropic (2023))45.6
Ifbench(Google Research (2023))31.5
Lcr(Artificial Analysis)21.0
Intelligence Index(Artificial Analysis)6.9
Terminalbench Hard(Stanford × Laude Institute (2026))3.8
Terminalbench V2 13.0
Coding Index(Artificial Analysis)2.9
Hle(Center for AI Safety + Scale AI (2025))2.6
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Chat60
Instruction Following60
Language60
Structured Output60
Legal50
Math50
Physics50
Reasoning50
Finance50
General50
Healthcare50
Agents50
Biology50
Tool Calling50
Chemistry40
Economics40
Long Context30
Multimodal30
Spatial Reasoning30
Communication30
Vision30
価格設定
入力価格無料
出力価格無料
混合価格(3:1)無料
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダーデータがありません