Qwen3.5 4B (Non-reasoning)
AlibabaQwenOpen WeightApache 2.0 · Commercial OK
Description
Qwen3.5-4B is a 4 billion parameter vision-language model using Gated DeltaNet hybrid architecture with a 3:1 ratio of linear attention to full softmax attention. It supports 262K native context length and delivers strong performance for its size across knowledge, reasoning, coding, and multilingual tasks.
Release Date
2026-03-02
Parameters
4.0B
Context Length
—
Modalities
—
Capability Radar
10
general
20
coding
71
reasoning
52
science
70
agents
50
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 107 | 28.0 | LS |
| Code Ranking | 452 | 27.0 | AA |
| General Ranking | 362 | 37.0 | AA |
| Science | 363 | 41.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
79.9%SR
VITA-Bench
22.0%SR
DeepPlanning
17.6%SR
Chat
IFEvalGoogle Research (2023)
89.8%SR
Multi-Challenge
49.0%SR
General
C-Eval
85.1%SR
MAXIFE
78.0%SR
Include
71.0%SR
NOVA-63
54.3%SR
Instruction Following
IFBench
59.2%SR
Language
MMLU-Redux
88.8%SR
MMLU-Pro
79.1%SR
MMMLU
76.1%SR
MMLU-ProX
71.5%SR
WMT24++
66.6%SR
Long Context
LongBench v2
50.0%SR
Math
HMMT25
76.8%SR
HMMT 2025
74.0%SR
PolyMATH
51.1%SR
Reasoning
Global PIQA
78.9%SR
GPQANYU + Cohere + Anthropic (2023)
76.2%SR
AA-LCR
57.0%SR
LiveCodeBench v6
55.8%SR
SuperGPQA
52.9%SR
Tool Calling
BFCL-V4
50.3%SR
AA Evaluation Indices
(Artificial Analysis)Tau2(Sierra + U Toronto + Vector Institute (2025))87.7
Gpqa(NYU + Cohere + Anthropic (2023))71.2
Lcr(Artificial Analysis)34.7
Ifbench(Google Research (2023))33.3
Terminalbench V2 121.3
Coding Index(Artificial Analysis)20.3
Terminalbench Hard(Stanford × Laude Institute (2026))11.4
Intelligence Index(Artificial Analysis)10.8
Hle(Center for AI Safety + Scale AI (2025))8.0
Tau Banking4.3
LLM Stats Category Scores
(LLM Stats (zeroeval))Language80
Biology80
Chat70
Instruction Following70
Legal70
Math70
Physics70
Structured Output70
Finance70
Healthcare70
Tool Calling70
Reasoning60
General60
Chemistry60
Long Context50
Multimodal50
Spatial Reasoning50
Communication50
Economics50
Vision50
Agents40
Pricing
Input Price$0.03 / 1M tokens
Output Price$0.15 / 1M tokens
Blended Price (3:1)$0.06 / 1M tokens
Speed
Tokens/sec20.6
Time to First Token0.72s
Time to Answer0.72s
Provider Price Ranking
Provider Price Ranking
1 providers
ProviderInputOutput
1AlibabaPRIMARY
$0.03
$0.15
Compare pricing across different API providers for this model.