Qwen3.5 2B (Reasoning)
AlibabaQwenओपन वेटApache 2.0 · व्यावसायिक उपयोग
विवरण
Qwen3.5-2B is a 2 billion parameter vision-language model using Gated DeltaNet hybrid architecture with a 3:1 ratio of linear attention to full softmax attention. It supports 262K native context length and features both thinking and non-thinking modes.
रिलीज़ तिथि
2026-03-02
पैरामीटर
2.0B
संदर्भ लंबाई
—
मोडैलिटीज़
—
क्षमता रडार
6
general
3
coding
46
reasoning
32
science
50
agents
30
multimodal
रैंकिंग
| डोमेन | #रैंक | स्कोर | स्रोत |
|---|---|---|---|
| एजेंटिक क्षमता | 106 | 31.0 | LS |
| कोडिंग रैंकिंग | 584 | 10.0 | AA |
| सामान्य रैंकिंग | 501 | 28.0 | AA |
| विज्ञान | 551 | 22.0 | AA |
बेंचमार्क स्कोर (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
48.8%स्वयं
Chat
IFEvalGoogle Research (2023)
78.6%स्वयं
Multi-Challenge
33.7%स्वयं
General
C-Eval
73.2%स्वयं
MAXIFE
60.6%स्वयं
Include
55.4%स्वयं
NOVA-63
46.4%स्वयं
Instruction Following
IFBench
41.3%स्वयं
Language
MMLU-Redux
79.6%स्वयं
MMLU-Pro
66.5%स्वयं
MMMLU
63.1%स्वयं
MMLU-ProX
52.3%स्वयं
WMT24++
45.8%स्वयं
Long Context
LongBench v2
38.7%स्वयं
Math
PolyMATH
26.1%स्वयं
Reasoning
Global PIQA
69.3%स्वयं
GPQANYU + Cohere + Anthropic (2023)
51.6%स्वयं
SuperGPQA
37.5%स्वयं
AA-LCR
25.6%स्वयं
Tool Calling
BFCL-V4
43.6%स्वयं
AA मूल्यांकन सूचकांक
(Artificial Analysis)Tau2(Sierra + U Toronto + Vector Institute (2025))69.0
Gpqa(NYU + Cohere + Anthropic (2023))45.6
Ifbench(Google Research (2023))31.5
Lcr(Artificial Analysis)21.0
Intelligence Index(Artificial Analysis)6.9
Terminalbench Hard(Stanford × Laude Institute (2026))3.8
Terminalbench V2 13.0
Coding Index(Artificial Analysis)2.9
Hle(Center for AI Safety + Scale AI (2025))2.6
LLM Stats श्रेणी स्कोर
(LLM Stats (zeroeval))Chat60
Instruction Following60
Language60
Structured Output60
Legal50
Math50
Physics50
Reasoning50
Finance50
General50
Healthcare50
Agents50
Biology50
Tool Calling50
Chemistry40
Economics40
Long Context30
Multimodal30
Spatial Reasoning30
Communication30
Vision30
मूल्य निर्धारण
इनपुट मूल्यमुफ्त
आउटपुट मूल्यमुफ्त
मिश्रित मूल्य (3:1)मुफ्त
गति
टोकन/सेकंड0.0
पहले टोकन में देरी0.00s
पहले उत्तर में देरी0.00s
प्रदाता मूल्य रैंकिंग
कोई प्रदाता डेटा उपलब्ध नहीं