Skip to main content

Qwen3.5 9B (Non-reasoning)

AlibabaQwenOpen WeightApache 2.0 · Commercial OK

Description

Qwen3.5-9B is a 9 billion parameter vision-language model using Gated DeltaNet hybrid architecture with a 3:1 ratio of linear attention to full softmax attention. It supports 262K native context length and delivers strong performance across knowledge, reasoning, coding, and multilingual tasks.

Release Date
2026-03-02
Parameters
9.0B
Context Length
262K
Modalities
image, text, video

Capability Radar

18
general
24
coding
79
reasoning
47
science
70
agents
60
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability123
35.0
LS
Code Ranking309
36.0
AA
General Ranking268
47.0
AA
Science275
47.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

t2-bench79.1%SR
BFCL-V466.1%SR
VITA-Bench29.8%SR
DeepPlanning18.0%SR

Biology

GPQANYU + Cohere + Anthropic (2023)81.7%SR

Chemistry

SuperGPQA58.2%SR

Communication

Multi-Challenge54.5%SR

Finance

MMLU-Pro82.5%SR
MMLU-ProX76.3%SR

General

IFEvalGoogle Research (2023)91.5%SR
MMLU-Redux91.1%SR
C-Eval88.2%SR
MAXIFE83.4%SR
Global PIQA83.2%SR
MMMLU81.2%SR
Include75.6%SR
LiveCodeBench v665.6%SR
IFBench64.5%SR
NOVA-6355.9%SR
LongBench v255.2%SR

Language

WMT24++72.6%SR

Long Context

AA-LCR63.0%SR

Math

HMMT 202583.2%SR
HMMT2582.9%SR
PolyMATH57.3%SR

AA Evaluation Indices

(Artificial Analysis)
Coding Index(Artificial Analysis)
23.5
Intelligence Index(Artificial Analysis)
20.6
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.9
Gpqa(NYU + Cohere + Anthropic (2023))
0.8
Lcr(Artificial Analysis)
0.5
Ifbench(Google Research (2023))
0.4
Scicode(UIUC + Argonne National Lab (2024))
0.3
Terminalbench V2 1
0.2
Terminalbench Hard(Stanford × Laude Institute (2026))
0.2
Hle(Center for AI Safety + Scale AI (2025))
0.1

LLM Stats Category Scores

(LLM Stats (zeroeval))
Math
80
Instruction Following
80
Language
80
Biology
80
Legal
70
Physics
70
Reasoning
70
Structured Output
70
Finance
70
General
70
Healthcare
70
Chemistry
70
Tool Calling
70
Long Context
60
Multimodal
60
Spatial Reasoning
60
Economics
60
Vision
60
Agents
50
Communication
50

Pricing

Input Price$0.17 / 1M tokens
Output Price$0.25 / 1M tokens
Blended Price (3:1)$0.19 / 1M tokens

Speed

Tokens/sec91.8
Time to First Token0.46s
Time to Answer0.46s

Provider Price Ranking

Provider Price Ranking

7 providers

Cheapest: NanoGPTMost Expensive: Alibaba
ProviderInputOutput
1NanoGPTCheapest
$0.05
$0.15
2Merge Gateway
$0.09
$0.13
3OpenRouter
$0.1
$0.15
4Kilo Gateway
$0.1
$0.15
5Mixlayer
$0.1
$0.4
6TensorX
$0.15
$0.2
7AlibabaPRIMARY
$0.17
$0.25

Compare pricing across different API providers for this model.

External Sources