Qwen3.8-Flash-Next
Description
Qwen3.8-Flash-Next is an open-weight experimental preview of the architecture planned for Qwen4, with Hybrid Attention (QSA), Gated Residual, and N-gram Embedding. The card reports 125B parameters with 6B activated, plus 51B n-gram embedding and 4B MTP; Hugging Face BF16 safetensors total 179,999,981,424 parameters (~180B stored). It is a causal LM with a vision encoder for text, image, and video, a native 262,144-token context extensible to 1,000,000 tokens, and thinking on by default (enable_thinking, preserve_thinking, reasoning_effort). This catalog entry is the open-weight checkpoint, not the separate production Qwen3.8-Flash API on Qwen Cloud.
Capability Radar
Science is estimated from LLM Stats science scores or reasoning when no dedicated science benchmarks are available.
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Multimodal Ranking | 25 | 62.0 | LS |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
Code
Instruction Following
Math
Multimodal
Reasoning
Vision
AA Evaluation Indices
(Artificial Analysis)No AA evaluation data available
LLM Stats Category Scores
(LLM Stats (zeroeval))Pricing
No pricing data available
Speed
No speed data available
Provider Price Ranking
No provider data available