Qwen3.5 4B (Non-reasoning)
AlibabaQwenओपन वेटApache 2.0 · व्यावसायिक उपयोग
विवरण
Qwen3.5-4B is a 4 billion parameter vision-language model using Gated DeltaNet hybrid architecture with a 3:1 ratio of linear attention to full softmax attention. It supports 262K native context length and delivers strong performance for its size across knowledge, reasoning, coding, and multilingual tasks.
रिलीज़ तिथि
2026-03-02
पैरामीटर
4.0B
संदर्भ लंबाई
—
मोडैलिटीज़
—
क्षमता रडार
10
general
20
coding
71
reasoning
52
science
70
agents
50
multimodal
रैंकिंग
| डोमेन | #रैंक | स्कोर | स्रोत |
|---|---|---|---|
| एजेंटिक क्षमता | 106 | 28.0 | LS |
| कोडिंग रैंकिंग | 443 | 27.0 | AA |
| सामान्य रैंकिंग | 356 | 37.0 | AA |
| विज्ञान | 355 | 41.0 | AA |
बेंचमार्क स्कोर (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
79.9%स्वयं
VITA-Bench
22.0%स्वयं
DeepPlanning
17.6%स्वयं
Chat
IFEvalGoogle Research (2023)
89.8%स्वयं
Multi-Challenge
49.0%स्वयं
General
C-Eval
85.1%स्वयं
MAXIFE
78.0%स्वयं
Include
71.0%स्वयं
NOVA-63
54.3%स्वयं
Instruction Following
IFBench
59.2%स्वयं
Language
MMLU-Redux
88.8%स्वयं
MMLU-Pro
79.1%स्वयं
MMMLU
76.1%स्वयं
MMLU-ProX
71.5%स्वयं
WMT24++
66.6%स्वयं
Long Context
LongBench v2
50.0%स्वयं
Math
HMMT25
76.8%स्वयं
HMMT 2025
74.0%स्वयं
PolyMATH
51.1%स्वयं
Reasoning
Global PIQA
78.9%स्वयं
GPQANYU + Cohere + Anthropic (2023)
76.2%स्वयं
AA-LCR
57.0%स्वयं
LiveCodeBench v6
55.8%स्वयं
SuperGPQA
52.9%स्वयं
Tool Calling
BFCL-V4
50.3%स्वयं
AA मूल्यांकन सूचकांक
(Artificial Analysis)Tau2(Sierra + U Toronto + Vector Institute (2025))87.7
Gpqa(NYU + Cohere + Anthropic (2023))71.2
Lcr(Artificial Analysis)34.7
Ifbench(Google Research (2023))33.3
Terminalbench V2 121.3
Coding Index(Artificial Analysis)20.3
Terminalbench Hard(Stanford × Laude Institute (2026))11.4
Intelligence Index(Artificial Analysis)10.8
Hle(Center for AI Safety + Scale AI (2025))8.0
Tau Banking4.3
LLM Stats श्रेणी स्कोर
(LLM Stats (zeroeval))Language80
Biology80
Chat70
Instruction Following70
Legal70
Math70
Physics70
Structured Output70
Finance70
Healthcare70
Tool Calling70
Reasoning60
General60
Chemistry60
Long Context50
Multimodal50
Spatial Reasoning50
Communication50
Economics50
Vision50
Agents40
मूल्य निर्धारण
इनपुट मूल्य$0.03 / 1M टोकन
आउटपुट मूल्य$0.15 / 1M टोकन
मिश्रित मूल्य (3:1)$0.06 / 1M टोकन
गति
टोकन/सेकंड25.9
पहले टोकन में देरी0.49s
पहले उत्तर में देरी0.49s
प्रदाता मूल्य रैंकिंग
प्रदाता मूल्य रैंकिंग
1 प्रदाता
प्रदाताइनपुटआउटपुट
1Alibabaप्राथमिक
$0.03
$0.15
इस मॉडल के लिए विभिन्न API प्रदाताओं के मूल्य निर्धारण की तुलना करें।