Qwen3.8-Flash-Next
AlibabaQwenOpen WeightQwen Community License 1.0 · Commercial OK
Description
Qwen3.8-Flash-Next is an open-weight experimental preview of the architecture planned for Qwen4, with Hybrid Attention (QSA), Gated Residual, and N-gram Embedding. The card reports 125B parameters with 6B activated, plus 51B n-gram embedding and 4B MTP; Hugging Face BF16 safetensors total 179,999,981,424 parameters (~180B stored). It is a causal LM with a vision encoder for text, image, and video, a native 262,144-token context extensible to 1,000,000 tokens, and thinking on by default (enable_thinking, preserve_thinking, reasoning_effort). This catalog entry is the open-weight checkpoint, not the separate production Qwen3.8-Flash API on Qwen Cloud.
Release Date
2026-08-26
Parameters
125.0B
Context Length
—
Modalities
—
Capability Radar
52
general
69
coding
92
reasoning
66
science
60
agents
70
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 18 | 93.0 | AA |
| General Ranking | 19 | 88.0 | AA |
| Math Reasoning | 29 | 86.0 | LB |
| Multimodal Ranking | 25 | 62.0 | LS |
| Reasoning | 17 | 87.0 | LB |
| Science | 54 | 81.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
AndroidWorld
84.5%SR
CoWorkBench
73.9%SR
Toolathlon
73.5%SR
ClawEval-MM
60.4%SR
Job Bench
55.7%SR
Agents' Last Exam
51.2%SR
RecreationBench
49.9%SR
Code
Vision2Web
64.0%SR
DeepSWE 1.1
58.7%SR
NL2Repo
48.1%SR
Instruction Following
IFBench
81.3%SR
Math
MathVision
95.7%SR
Multimodal
OSWorld 2.0
19.4%SR
Reasoning
LiveCodeBench v6
91.9%SR
GPQANYU + Cohere + Anthropic (2023)
91.7%SR
CharXiv-R
90.6%SR
SWE-bench Multilingual
81.0%SR
SWE-Bench ProPrinceton NLP (2024)
62.5%SR
Humanity's Last Exam
35.9%SR
Vision
RealWorldQA
88.5%SR
LVBench
76.6%SR
ERQA
72.3%SR
AA Evaluation Indices
(Artificial Analysis)Gpqa(NYU + Cohere + Anthropic (2023))92.3
Terminalbench V2 186.1
Lcr(Artificial Analysis)77.0
Coding Index(Artificial Analysis)73.1
Intelligence Index(Artificial Analysis)55.8
Scicode(UIUC + Argonne National Lab (2024))46.9
Tau Banking45.4
Hle(Center for AI Safety + Scale AI (2025))38.0
LLM Stats Category Scores
(LLM Stats (zeroeval))Physics90
Biology90
Chemistry90
Instruction Following80
Long Context80
Spatial Reasoning80
Math70
Multimodal70
Reasoning70
General70
Vision70
Productivity60
Agents60
Code60
Tool Calling60
Pricing
Input PriceFree
Output PriceFree
Blended Price (3:1)Free
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
No provider data available