Qwen3.8 Max (0902)
AlibabaQwenオープンウエイトQwen3.8-Max License · 商用利用可
説明
Qwen3.8 Max, released as the Qwen3.8-2.4T-A95B open-weight checkpoint, is Qwen's flagship mixture-of-experts model for coding, research, professional workflows, and long-horizon agents. It has 2.4 trillion total parameters with 95 billion active parameters, a native 262,144-token context window extensible to about 1 million tokens, and always-on thinking. QwenCloud serves the same model in the Qwen3.8 Max API family with text and image input.
リリース日
2026-09-02
パラメータ
2.4T
コンテキスト長
1.0M
モダリティ
image, pdf, text, video
能力レーダー
45
general
72
coding
93
reasoning
69
science
60
agents
90
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| エージェント能力 | 18 | 61.0 | LS |
| コーディングランキング | 43 | 92.0 | AA |
| 総合ランキング | 36 | 78.0 | AA |
| 数学的推論 | 31 | 91.0 | LB |
| マルチモーダルランキング | 1 | 84.0 | LS |
| 推論 | 24 | 88.0 | LB |
| 科学 | 62 | 80.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Agents
QwenSVG
1713.00 / 2000自己申告
AndroidWorld
85.3%自己申告
MobileWorld
77.8%自己申告
AndroidBench
75.1%自己申告
CoWorkBench
74.8%自己申告
Toolathlon
72.5%自己申告
Workspace Bench
67.7%自己申告
Job Bench
53.4%自己申告
OneMillion Bench
52.5%自己申告
Agents' Last Exam
52.4%自己申告
MLS-Bench Lite
41.0%自己申告
AutomationBench
27.3%自己申告
Code
QwenReactBench
1724.00 / 2000自己申告
PaperBench
93.0%自己申告
QwenSWEBench
80.7%自己申告
FrontierSWE
73.5%自己申告
SkillsBench
70.2%自己申告
Vision2Web
69.0%自己申告
QwenQoderBench
58.4%自己申告
DeepSWE 1.1
56.6%自己申告
NL2Repo
55.9%自己申告
Healthcare
HealthBench
60.2%自己申告
Instruction Following
IFBench
82.8%自己申告
Knowledge
PLawBench
73.2%自己申告
PRBench-Finance
58.3%自己申告
PRBench-Legal
57.6%自己申告
Long Context
MRCR v2 (8-needle)
92.9%自己申告
LongBench v2
66.3%自己申告
Multimodal
OSWorld-Verified
86.1%自己申告
Reasoning
GPQANYU + Cohere + Anthropic (2023)
92.6%自己申告
Terminal-Bench 2.1
86.6%自己申告
SWE-Bench ProPrinceton NLP (2024)
67.7%自己申告
Humanity's Last Exam (with tools, text-only)
56.2%自己申告
Humanity's Last Exam
43.6%自己申告
Search
WideSearch
81.9%自己申告
Vision
VideoMME w sub.
90.4%自己申告
RealWorldQA
88.0%自己申告
ScreenSpot Pro
84.5%自己申告
MMMU-Pro
82.3%自己申告
LVBench
81.8%自己申告
ERQA
77.8%自己申告
PerceptionBench
63.5%自己申告
AA評価指数
(Artificial Analysis)Gpqa(NYU + Cohere + Anthropic (2023))92.8
Terminalbench V2 188.8
Lcr(Artificial Analysis)80.3
Coding Index(Artificial Analysis)76.2
Scicode(UIUC + Argonne National Lab (2024))52.1
Tau Banking47.8
Intelligence Index(Artificial Analysis)45.4
Hle(Center for AI Safety + Scale AI (2025))43.1
Terminalbench V4 038.9
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Physics90
Biology90
Chemistry90
Video90
Instruction Following80
Multimodal80
Search80
Spatial Reasoning80
Grounding80
Vision80
Long Context70
Reasoning70
Structured Output70
General70
Agents70
Code70
Productivity60
Healthcare60
Tool Calling60
Math50
価格設定
入力価格$2 / 1Mトークン
出力価格$6 / 1Mトークン
混合価格(3:1)$3 / 1Mトークン
キャッシュ読み取り価格$0.25 / 1Mトークン
キャッシュ書き込み価格$2.5 / 1Mトークン
速度
トークン/秒39.8
初トークン遅延1.81s
初回答遅延52.03s
プロバイダー価格ランキング
プロバイダー価格ランキング
6 プロバイダー
最安: DeepInfra最高: EmpirioLabs AI
プロバイダー入力出力
1DeepInfra最安
$0
$0
2Novita
$0
$0.00001
3Fireworks
$0
$0.00001
4Together
$0
$0.00001
5Alibabaプライマリ
$2
$6
6EmpirioLabs AI
$2
$6
このモデルの異なるAPIプロバイダー間の価格を比較。