Qwen3.5 35B A3B (Non-reasoning)
AlibabaQwenOpen WeightApache 2.0 · Commercial OK
Description
Qwen3.5-35B-A3B is a multimodal Mixture-of-Experts model with 35 billion total parameters and 3 billion activated parameters. It combines strong reasoning, coding, agentic, and visual understanding performance with production-friendly efficiency and a native 262K context window.
Release Date
2026-02-24
Parameters
35.0B
Context Length
262K
Modalities
audio, image, text, video
Capability Radar
15
general
37
coding
82
reasoning
61
science
60
agents
80
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 62 | 46.0 | LS |
| Code Ranking | 340 | 44.0 | AA |
| General Ranking | 287 | 44.0 | AA |
| Multimodal Ranking | 102 | 46.0 | LS |
| Science | 248 | 52.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
81.2%SR
VITA-Bench
31.9%SR
DeepPlanning
22.8%SR
Chat
IFEvalGoogle Research (2023)
91.9%SR
Multi-Challenge
60.0%SR
Code
FullStackBench en
58.1%SR
FullStackBench zh
55.0%SR
General
C-Eval
90.2%SR
MAXIFE
86.6%SR
Include
79.7%SR
AndroidWorld_SR
71.1%SR
NOVA-63
57.1%SR
Healthcare
PMC-VQA
62.0%SR
MedXpertQA
61.4%SR
Instruction Following
IFBench
70.2%SR
Language
MMLU-Redux
93.3%SR
MMLU-Pro
85.3%SR
MMMLU
85.2%SR
MMLU-ProX
81.0%SR
WMT24++
76.3%SR
Long Context
LongBench v2
59.0%SR
Math
HMMT25
89.2%SR
HMMT 2025
89.0%SR
MathVista-Mini
86.2%SR
DynaMath
85.0%SR
MathVision
83.9%SR
CodeForces
0.82 / 3000SR
PolyMATH
64.4%SR
Multimodal
VideoMME w/o sub.
82.5%SR
MMMU
81.4%SR
VideoMMMU
80.4%SR
TIR-Bench
55.5%SR
OSWorld-Verified
54.5%SR
Reasoning
Global PIQA
86.6%SR
GPQANYU + Cohere + Anthropic (2023)
84.2%SR
CharXiv-R
77.5%SR
LiveCodeBench v6
74.6%SR
BrowseComp-zh
69.5%SR
SWE-Bench Verified
69.2%SR
SuperGPQA
63.4%SR
BrowseCompOpenAI (2025)
61.0%SR
AA-LCR
58.5%SR
Humanity's Last Exam
47.4%SR
Seal-0
41.4%SR
Terminal-Bench 2.0Stanford × Laude Institute (2026)
40.5%SR
OJBench
36.0%SR
Search
WideSearch
57.1%SR
Tool Calling
BFCL-V4
67.3%SR
Video
MLVU
85.6%SR
Vision
CountBench
0.98 / 100SR
VLMsAreBlind
97.0%SR
V*
92.7%SR
AI2D
92.6%SR
MMBench-V1.1
91.5%SR
OCRBench
91.0%SR
OmniDocBench 1.5
89.3%SR
RefCOCO-avg
0.89 / 100SR
VideoMME w sub.
86.6%SR
RealWorldQA
84.1%SR
EmbSpatialBench
0.83 / 100SR
MMStar
81.9%SR
CC-OCR
80.7%SR
LingoQA
79.2%SR
SlakeVQA
78.7%SR
MMMU-Pro
75.1%SR
MVBench
74.8%SR
MMVU
72.3%SR
LVBench
71.4%SR
ScreenSpot Pro
68.6%SR
Hallusion Bench
67.9%SR
ERQA
64.8%SR
RefSpatialBench
0.64 / 100SR
MMLongBench-Doc
0.59 / 100SR
SimpleVQA
0.58 / 100SR
ODinW
42.6%SR
BabyVision
38.4%SR
ZEROBench-Sub
0.34 / 100SR
SUNRGBD
0.33 / 100SR
Nuscene
14.6%SR
Hypersim
0.13 / 100SR
ZEROBench
0.08 / 100SR
AA Evaluation Indices
(Artificial Analysis)Tau2(Sierra + U Toronto + Vector Institute (2025))86.3
Gpqa(NYU + Cohere + Anthropic (2023))81.9
Lcr(Artificial Analysis)63.0
Ifbench(Google Research (2023))44.5
Terminalbench V2 140.8
Coding Index(Artificial Analysis)37.0
Intelligence Index(Artificial Analysis)15.1
Hle(Center for AI Safety + Scale AI (2025))13.4
Terminalbench Hard(Stanford × Laude Institute (2026))10.6
Tau Banking4.9
LLM Stats Category Scores
(LLM Stats (zeroeval))Chat80
Image To Text80
Instruction Following80
Language80
Legal80
Math80
Physics80
Structured Output80
Embodied80
Finance80
Biology80
Text-to-image80
Video80
Long Context70
Multimodal70
Reasoning70
Spatial Reasoning70
Frontend Development70
General70
Grounding70
Healthcare70
Chemistry70
Vision70
Search60
Code60
Communication60
Economics60
Tool Calling60
Agents50
3d20
Spatial10
Pricing
Input Price$0.25 / 1M tokens
Output Price$2 / 1M tokens
Blended Price (3:1)$0.688 / 1M tokens
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
Provider Price Ranking
6 providers
Cheapest: DeepInfraMost Expensive: DevPass (LLM Gateway)
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2AIHubMix
$0.0564
$0.4512
3302.AI
$0.06
$0.46
4Requesty
$0.14
$1
5AlibabaPRIMARY
$0.25
$2
6DevPass (LLM Gateway)
$0.25
$2
Compare pricing across different API providers for this model.