Qwen3.8 Max (0902)
AlibabaQwen开源权重Qwen3.8-Max License · 商用许可
描述
Qwen3.8 Max, released as the Qwen3.8-2.4T-A95B open-weight checkpoint, is Qwen's flagship mixture-of-experts model for coding, research, professional workflows, and long-horizon agents. It has 2.4 trillion total parameters with 95 billion active parameters, a native 262,144-token context window extensible to about 1 million tokens, and always-on thinking. QwenCloud serves the same model in the Qwen3.8 Max API family with text and image input.
发布日期
2026-09-02
参数规模
2.4T
上下文长度
1.0M
支持模态
image, pdf, text, video
能力雷达图
45
general
72
coding
93
reasoning
69
science
60
agents
90
multimodal
排行榜排名
基准测试分数 (LLM Stats)
(LLM Stats (zeroeval))Agents
QwenSVG
1713.00 / 2000自报
AndroidWorld
85.3%自报
MobileWorld
77.8%自报
AndroidBench
75.1%自报
CoWorkBench
74.8%自报
Toolathlon
72.5%自报
Workspace Bench
67.7%自报
Job Bench
53.4%自报
OneMillion Bench
52.5%自报
Agents' Last Exam
52.4%自报
MLS-Bench Lite
41.0%自报
AutomationBench
27.3%自报
Code
QwenReactBench
1724.00 / 2000自报
PaperBench
93.0%自报
QwenSWEBench
80.7%自报
FrontierSWE
73.5%自报
SkillsBench
70.2%自报
Vision2Web
69.0%自报
QwenQoderBench
58.4%自报
DeepSWE 1.1
56.6%自报
NL2Repo
55.9%自报
Healthcare
HealthBench
60.2%自报
Instruction Following
IFBench
82.8%自报
Knowledge
PLawBench
73.2%自报
PRBench-Finance
58.3%自报
PRBench-Legal
57.6%自报
Long Context
MRCR v2 (8-needle)
92.9%自报
LongBench v2
66.3%自报
Multimodal
OSWorld-Verified
86.1%自报
Reasoning
GPQANYU + Cohere + Anthropic (2023)
92.6%自报
Terminal-Bench 2.1
86.6%自报
SWE-Bench ProPrinceton NLP (2024)
67.7%自报
Humanity's Last Exam (with tools, text-only)
56.2%自报
Humanity's Last Exam
43.6%自报
Search
WideSearch
81.9%自报
Vision
VideoMME w sub.
90.4%自报
RealWorldQA
88.0%自报
ScreenSpot Pro
84.5%自报
MMMU-Pro
82.3%自报
LVBench
81.8%自报
ERQA
77.8%自报
PerceptionBench
63.5%自报
AA 评测指数
(Artificial Analysis)Gpqa(NYU + Cohere + Anthropic (2023))92.8
Terminalbench V2 188.8
Lcr(Artificial Analysis)80.3
Coding Index(Artificial Analysis)76.2
Scicode(UIUC + Argonne National Lab (2024))52.1
Tau Banking47.8
Intelligence Index(Artificial Analysis)45.4
Hle(Center for AI Safety + Scale AI (2025))43.1
Terminalbench V4 038.9
LLM Stats 分类评分
(LLM Stats (zeroeval))Physics90
Biology90
Chemistry90
Video90
Instruction Following80
Multimodal80
Search80
Spatial Reasoning80
Grounding80
Vision80
Long Context70
Reasoning70
Structured Output70
General70
Agents70
Code70
Productivity60
Healthcare60
Tool Calling60
Math50
定价
输入价格$2 / 1M tokens
输出价格$6 / 1M tokens
混合价格(3:1)$3 / 1M tokens
缓存读取价格$0.25 / 1M tokens
缓存写入价格$2.5 / 1M tokens
速度
Tokens/秒39.3
首Token延迟1.77s
首回答延迟52.69s
供应商价格排行
供应商价格排行
6 个供应商
最便宜: DeepInfra最贵: EmpirioLabs AI
供应商输入输出
1DeepInfra最便宜
$0
$0
2Novita
$0
$0.00001
3Fireworks
$0
$0.00001
4Together
$0
$0.00001
5Alibaba主要
$2
$6
6EmpirioLabs AI
$2
$6
比较该模型在不同 API 供应商之间的定价。