Qwen3.5 2B (Reasoning)
AlibabaQwen開源權重Apache 2.0 · 商用許可
描述
Qwen3.5-2B is a 2 billion parameter vision-language model using Gated DeltaNet hybrid architecture with a 3:1 ratio of linear attention to full softmax attention. It supports 262K native context length and features both thinking and non-thinking modes.
發布日期
2026-03-02
參數規模
2.0B
上下文長度
—
支援模態
—
能力雷達圖
6
general
3
coding
46
reasoning
22
science
50
agents
30
multimodal
排行榜排名
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
48.8%自報
BFCL-V4
43.6%自報
Biology
GPQANYU + Cohere + Anthropic (2023)
51.6%自報
Chemistry
SuperGPQA
37.5%自報
Communication
Multi-Challenge
33.7%自報
Finance
MMLU-Pro
66.5%自報
MMLU-ProX
52.3%自報
General
MMLU-Redux
79.6%自報
IFEvalGoogle Research (2023)
78.6%自報
C-Eval
73.2%自報
Global PIQA
69.3%自報
MMMLU
63.1%自報
MAXIFE
60.6%自報
Include
55.4%自報
NOVA-63
46.4%自報
IFBench
41.3%自報
LongBench v2
38.7%自報
Language
WMT24++
45.8%自報
Long Context
AA-LCR
25.6%自報
Math
PolyMATH
26.1%自報
AA 評測指數
(Artificial Analysis)Intelligence Index(Artificial Analysis)7.4
Coding Index(Artificial Analysis)2.9
Tau2(Sierra + U Toronto + Vector Institute (2025))0.7
Gpqa(NYU + Cohere + Anthropic (2023))0.5
Ifbench(Google Research (2023))0.3
Lcr(Artificial Analysis)0.3
Terminalbench Hard(Stanford × Laude Institute (2026))0.0
Terminalbench V2 10.0
Scicode(UIUC + Argonne National Lab (2024))0.0
Hle(Center for AI Safety + Scale AI (2025))0.0
LLM Stats 分類評分
(LLM Stats (zeroeval))Structured Output60
Instruction Following60
Language60
Legal50
Math50
Physics50
Reasoning50
Finance50
General50
Healthcare50
Agents50
Biology50
Tool Calling50
Chemistry40
Economics40
Long Context30
Multimodal30
Spatial Reasoning30
Communication30
Vision30
定價
輸入價格免費
輸出價格免費
混合價格(3:1)免費
速度
Tokens/秒0.0
首Token延遲0.00s
首回答延遲0.00s
供應商價格排行
暫無提供商資料