跳轉到主要內容

Qwen3.8-Flash-Next

Alibaba Cloud / Qwen TeamQwen開源權重Qwen Community License 1.0 · 商用許可

描述

Qwen3.8-Flash-Next is an open-weight experimental preview of the architecture planned for Qwen4, with Hybrid Attention (QSA), Gated Residual, and N-gram Embedding. The card reports 125B parameters with 6B activated, plus 51B n-gram embedding and 4B MTP; Hugging Face BF16 safetensors total 179,999,981,424 parameters (~180B stored). It is a causal LM with a vision encoder for text, image, and video, a native 262,144-token context extensible to 1,000,000 tokens, and thinking on by default (enable_thinking, preserve_thinking, reasoning_effort). This catalog entry is the open-weight checkpoint, not the separate production Qwen3.8-Flash API on Qwen Cloud.

發布日期
2026-08-26
參數規模
125.0B
上下文長度
支援模態

能力雷達圖

70
general
60
coding
70
reasoning
77
science估算
60
agents
70
multimodal

缺少專門科學評測時,Science 由 LLM Stats 科學得分或推理能力估算。

排行榜排名

領域#排名分數來源
多模態榜25
62.0
LS

基準測試分數 (LLM Stats)

(LLM Stats (zeroeval))

Agents

AndroidWorld84.5%自報
CoWorkBench73.9%自報
Toolathlon73.5%自報
ClawEval-MM60.4%自報
Job Bench55.7%自報
Agents' Last Exam51.2%自報
RecreationBench49.9%自報

Code

Vision2Web64.0%自報
DeepSWE 1.158.7%自報
NL2Repo48.1%自報

Instruction Following

IFBench81.3%自報

Math

MathVision95.7%自報

Multimodal

OSWorld 2.019.4%自報

Reasoning

LiveCodeBench v691.9%自報
GPQANYU + Cohere + Anthropic (2023)91.7%自報
CharXiv-R90.6%自報
SWE-bench Multilingual81.0%自報
SWE-Bench ProPrinceton NLP (2024)62.5%自報
Humanity's Last Exam35.9%自報

Vision

RealWorldQA88.5%自報
LVBench76.6%自報
ERQA72.3%自報

AA 評測指數

(Artificial Analysis)

暫無 AA 評測資料

LLM Stats 分類評分

(LLM Stats (zeroeval))
Physics
90
Biology
90
Chemistry
90
Instruction Following
80
Long Context
80
Spatial Reasoning
80
Math
70
Multimodal
70
Reasoning
70
General
70
Vision
70
Productivity
60
Agents
60
Code
60
Tool Calling
60

定價

暫無定價資料

速度

暫無速度資料

供應商價格排行

暫無提供商資料

外部連結