Skip to main content

Qwen3.5 9B (Non-reasoning)

AlibabaQwenOpen WeightApache 2.0 · Commercial OK

Description

Qwen3.5-9B is a 9 billion parameter vision-language model using Gated DeltaNet hybrid architecture with a 3:1 ratio of linear attention to full softmax attention. It supports 262K native context length and delivers strong performance across knowledge, reasoning, coding, and multilingual tasks.

Release Date
2026-03-02
Parameters
9.0B
Context Length
262K
Modalities
image, text, video

Capability Radar

12
general
24
coding
79
reasoning
57
science
70
agents
60
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability100
35.0
LS
Code Ranking391
35.0
AA
General Ranking343
40.0
AA
Science299
47.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

t2-bench79.1%SR
VITA-Bench29.8%SR
DeepPlanning18.0%SR

Chat

IFEvalGoogle Research (2023)91.5%SR
Multi-Challenge54.5%SR

General

C-Eval88.2%SR
MAXIFE83.4%SR
Include75.6%SR
NOVA-6355.9%SR

Instruction Following

IFBench64.5%SR

Language

MMLU-Redux91.1%SR
MMLU-Pro82.5%SR
MMMLU81.2%SR
MMLU-ProX76.3%SR
WMT24++72.6%SR

Long Context

LongBench v255.2%SR

Math

HMMT 202583.2%SR
HMMT2582.9%SR
PolyMATH57.3%SR

Reasoning

Global PIQA83.2%SR
GPQANYU + Cohere + Anthropic (2023)81.7%SR
LiveCodeBench v665.6%SR
AA-LCR63.0%SR
SuperGPQA58.2%SR

Tool Calling

BFCL-V466.1%SR

AA Evaluation Indices

(Artificial Analysis)
Tau2(Sierra + U Toronto + Vector Institute (2025))
85.1
Gpqa(NYU + Cohere + Anthropic (2023))
78.6
Lcr(Artificial Analysis)
46.0
Ifbench(Google Research (2023))
37.8
Coding Index(Artificial Analysis)
23.5
Terminalbench V2 1
21.3
Terminalbench Hard(Stanford × Laude Institute (2026))
18.2
Intelligence Index(Artificial Analysis)
13.3
Hle(Center for AI Safety + Scale AI (2025))
9.4

LLM Stats Category Scores

(LLM Stats (zeroeval))
Instruction Following
80
Language
80
Math
80
Biology
80
Chat
70
Legal
70
Physics
70
Reasoning
70
Structured Output
70
Finance
70
General
70
Healthcare
70
Chemistry
70
Tool Calling
70
Long Context
60
Multimodal
60
Spatial Reasoning
60
Economics
60
Vision
60
Agents
50
Communication
50

Pricing

Input Price$0.17 / 1M tokens
Output Price$0.25 / 1M tokens
Blended Price (3:1)$0.19 / 1M tokens

Speed

Tokens/sec81.4
Time to First Token0.43s
Time to Answer0.43s

Provider Price Ranking

Provider Price Ranking

2 providers

Cheapest: DeepInfraMost Expensive: Alibaba
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2AlibabaPRIMARY
$0.17
$0.25

Compare pricing across different API providers for this model.

External Sources