Qwen3.5 4B (Non-reasoning)
AlibabaQwen开源权重Apache 2.0 · 商用许可
描述
Qwen3.5-4B is a 4 billion parameter vision-language model using Gated DeltaNet hybrid architecture with a 3:1 ratio of linear attention to full softmax attention. It supports 262K native context length and delivers strong performance for its size across knowledge, reasoning, coding, and multilingual tasks.
发布日期
2026-03-02
参数规模
4.0B
上下文长度
—
支持模态
—
能力雷达图
10
general
20
coding
71
reasoning
52
science
70
agents
50
multimodal
排行榜排名
基准测试分数 (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
79.9%自报
VITA-Bench
22.0%自报
DeepPlanning
17.6%自报
Chat
IFEvalGoogle Research (2023)
89.8%自报
Multi-Challenge
49.0%自报
General
C-Eval
85.1%自报
MAXIFE
78.0%自报
Include
71.0%自报
NOVA-63
54.3%自报
Instruction Following
IFBench
59.2%自报
Language
MMLU-Redux
88.8%自报
MMLU-Pro
79.1%自报
MMMLU
76.1%自报
MMLU-ProX
71.5%自报
WMT24++
66.6%自报
Long Context
LongBench v2
50.0%自报
Math
HMMT25
76.8%自报
HMMT 2025
74.0%自报
PolyMATH
51.1%自报
Reasoning
Global PIQA
78.9%自报
GPQANYU + Cohere + Anthropic (2023)
76.2%自报
AA-LCR
57.0%自报
LiveCodeBench v6
55.8%自报
SuperGPQA
52.9%自报
Tool Calling
BFCL-V4
50.3%自报
AA 评测指数
(Artificial Analysis)Tau2(Sierra + U Toronto + Vector Institute (2025))87.7
Gpqa(NYU + Cohere + Anthropic (2023))71.2
Lcr(Artificial Analysis)34.7
Ifbench(Google Research (2023))33.3
Terminalbench V2 121.3
Coding Index(Artificial Analysis)20.3
Terminalbench Hard(Stanford × Laude Institute (2026))11.4
Intelligence Index(Artificial Analysis)10.8
Hle(Center for AI Safety + Scale AI (2025))8.0
Tau Banking4.3
LLM Stats 分类评分
(LLM Stats (zeroeval))Language80
Biology80
Chat70
Instruction Following70
Legal70
Math70
Physics70
Structured Output70
Finance70
Healthcare70
Tool Calling70
Reasoning60
General60
Chemistry60
Long Context50
Multimodal50
Spatial Reasoning50
Communication50
Economics50
Vision50
Agents40
定价
输入价格$0.03 / 1M tokens
输出价格$0.15 / 1M tokens
混合价格(3:1)$0.06 / 1M tokens
速度
Tokens/秒25.9
首Token延迟0.53s
首回答延迟0.53s
供应商价格排行
供应商价格排行
1 个供应商
供应商输入输出
1Alibaba主要
$0.03
$0.15
比较该模型在不同 API 供应商之间的定价。