Skip to main content

Qwen3.8 27B (Xhigh)

AlibabaQwenOpen WeightApache 2.0 · Commercial OK

Description

Qwen3.8-27B is a dense 27.78-billion-parameter multimodal foundation model for coding, professional work, research, and long-horizon agents. It natively understands text, images, and video, supports configurable thinking effort and preserved reasoning across turns, and uses a 64-layer hybrid architecture with repeating Gated DeltaNet and full-attention blocks. Its native context window is 262,144 tokens and can be extended to approximately 1 million tokens. The open weights are released under Apache 2.0 and support Transformers, vLLM, SGLang, and TokenSpeed.

Release Date
2026-08-14
Parameters
27.8B
Context Length
1.0M
Modalities
image, text, video

Capability Radar

34
general
65
coding
91
reasoning
64
science
60
agents
70
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability64
45.0
LS
Code Ranking90
86.0
AA
General Ranking186
56.0
AA
Multimodal Ranking6
72.0
LS
Science117
71.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

AndroidWorld81.9%SR
CoWorkBench70.7%SR
WebArena-Verified64.8%SR
ClawEval-MM56.9%SR
RecreationBench47.1%SR
Agents' Last Exam42.9%SR
Job Bench33.4%SR

Code

QwenSWEBench79.0%SR
Vision2Web62.9%SR
NL2Repo42.3%SR
DeepSWE 1.142.2%SR
SWE-MM38.6%SR

Instruction Following

IFBench79.5%SR

Math

MathVision94.6%SR

Multimodal

OSWorld-Verified84.3%SR

Reasoning

LiveCodeBench v690.3%SR
CharXiv-R90.2%SR
GPQANYU + Cohere + Anthropic (2023)89.2%SR
Terminal-Bench 2.173.0%SR
SWE-Bench ProPrinceton NLP (2024)61.7%SR
SWE-Bench Multimodal38.6%SR
Humanity's Last Exam30.8%SR

Vision

OmniDocBench 1.591.1%SR
RealWorldQA85.9%SR
BabyVision85.6%SR
ERQA65.5%SR

AA Evaluation Indices

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
90.5
Lcr(Artificial Analysis)
82.0
Terminalbench V2 1
79.8
Coding Index(Artificial Analysis)
68.1
Tau Banking
48.0
Scicode(UIUC + Argonne National Lab (2024))
46.6
Hle(Center for AI Safety + Scale AI (2025))
33.9
Intelligence Index(Artificial Analysis)
33.7
Terminalbench V4 0
5.6

LLM Stats Category Scores

(LLM Stats (zeroeval))
Physics
90
Structured Output
90
Biology
90
Chemistry
90
Instruction Following
80
Spatial Reasoning
80
Multimodal
70
Reasoning
70
General
70
Vision
70
Math
60
Agents
60
Tool Calling
60
Productivity
50
Code
50

Pricing

Input Price$0.5 / 1M tokens
Output Price$3 / 1M tokens
Blended Price (3:1)$1.125 / 1M tokens
Cache Read Price$0.085 / 1M tokens
Cache Write Price$0.53125 / 1M tokens

Speed

Tokens/sec46.8
Time to First Token1.21s
Time to Answer43.94s

Provider Price Ranking

Provider Price Ranking

3 providers

Cheapest: DeepInfraMost Expensive: Alibaba
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2EmpirioLabs AI
$0.17
$0.5
3AlibabaPRIMARY
$0.5
$3

Compare pricing across different API providers for this model.

External Sources