Skip to main content

Qwen3.8 Flash

Alibaba Cloud / Qwen TeamQwenProprietary

Description

Qwen3.8 Flash is the production QwenCloud / OpenRouter API model (id qwen3.8-flash), not the open-weight Qwen3.8-Flash-Next checkpoint. Hugging Face states Flash is the official managed version based on Flash-Next with production features such as 1M context by default and official built-in tools. Architecture inherits Flash-Next (125B parameters with 6B activated, plus 51B n-gram embedding and 4B MTP). Multimodal text, image, and video understanding; thinking on by default; function calling, built-in tools, and structured output. QwenCloud lists 1M context, ~128k max output, and ~256k thinking budget. Open weights remain under the separate catalog card qwen3.8-flash-next.

Release Date
2026-08-26
Parameters
125.0B
Context Length
1.0M
Modalities
image, text, video

Capability Radar

70
general
60
coding
70
reasoning
77
scienceest.
60
agents
70
multimodal

Science is estimated from LLM Stats science scores or reasoning when no dedicated science benchmarks are available.

Rankings

Domain#RankScoreSource
Agentic Capability4
67.0
LS
Multimodal Ranking33
60.0
LS

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

AndroidWorld84.5%SR
CoWorkBench73.9%SR
Toolathlon73.5%SR
ClawEval-MM60.4%SR
Job Bench55.7%SR
Agents' Last Exam51.2%SR
RecreationBench49.9%SR

Code

Vision2Web64.0%SR
DeepSWE 1.158.7%SR
NL2Repo48.1%SR

Instruction Following

IFBench81.3%SR

Math

MathVision95.7%SR

Multimodal

OSWorld 2.019.4%SR

Reasoning

LiveCodeBench v691.9%SR
GPQANYU + Cohere + Anthropic (2023)91.7%SR
CharXiv-R90.6%SR
SWE-bench Multilingual81.0%SR
SWE-Bench ProPrinceton NLP (2024)62.5%SR
Humanity's Last Exam35.9%SR

Vision

RealWorldQA88.5%SR
LVBench76.6%SR
ERQA72.3%SR

AA Evaluation Indices

(Artificial Analysis)

No AA evaluation data available

LLM Stats Category Scores

(LLM Stats (zeroeval))
Physics
90
Biology
90
Chemistry
90
Long Context
80
Spatial Reasoning
80
Instruction Following
80
Math
70
Multimodal
70
Reasoning
70
General
70
Vision
70
Productivity
60
Agents
60
Code
60
Tool Calling
60

Pricing

Input Price$0 / 1M tokens
Output Price$0 / 1M tokens
Blended Price (3:1)$0 / 1M tokens
Cache Read Price$0.016 / 1M tokens
Cache Write Price$0.2 / 1M tokens

Speed

No speed data available

Provider Price Ranking

Provider Price Ranking

12 providers

Cheapest: Alibaba Cloud / Qwen TeamMost Expensive: Eden AI
ProviderInputOutput
1Alibaba Cloud / Qwen TeamPRIMARY
$0
$0
2Novita
$0
$0
3Alibaba (China)
$0.11875
$0.40073
4Vancine
$0.12
$0.38
5CrossModel
$0.13
$0.43
6Alibaba
$0.15
$0.47
7OpenRouter
$0.15
$0.47
8OpenCode Go
$0.15
$0.47
9Kilo Gateway
$0.15
$0.47
10DevPass (LLM Gateway)
$0.15
$0.47
11Charm Hyper
$0.15
$0.47
12Eden AI
$0.16
$0.47

Compare pricing across different API providers for this model.

External Sources