Qwen3.8 Max
AlibabaQwenオープンウエイトQwen3.8-Max License · 商用利用可
説明
Qwen3.8 Max, released as the Qwen3.8-2.4T-A95B open-weight checkpoint, is Qwen's flagship mixture-of-experts model for coding, research, professional workflows, and long-horizon agents. It has 2.4 trillion total parameters with 95 billion active parameters, a native 262,144-token context window extensible to about 1 million tokens, and always-on thinking. QwenCloud serves the same model in the Qwen3.8 Max API family with text and image input.
リリース日
2026-08-03
パラメータ
2.4T
コンテキスト長
1.0M
モダリティ
image, pdf, text, video
能力レーダー
55
general
69
coding
93
reasoning
69
science
60
agents
90
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| エージェント能力 | 4 | 79.0 | LS |
| コーディングランキング | 26 | 91.0 | AA |
| 総合ランキング | 8 | 92.0 | AA |
| 数学的推論 | 17 | 91.0 | LB |
| マルチモーダルランキング | 10 | 65.0 | LS |
| 推論 | 13 | 88.0 | LB |
| 科学 | 26 | 87.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Agents
QwenReactBench
1724.00 / 2000自己申告
QwenSVG
1713.00 / 2000自己申告
PaperBench
93.0%自己申告
Terminal-Bench 2.1
86.6%自己申告
OSWorld-Verified
86.1%自己申告
AndroidWorld
85.3%自己申告
WideSearch
81.9%自己申告
QwenSWEBench
80.7%自己申告
MobileWorld
77.8%自己申告
AndroidBench
75.1%自己申告
CoWorkBench
74.8%自己申告
FrontierSWE
73.5%自己申告
Toolathlon
72.5%自己申告
SkillsBench
70.2%自己申告
Workspace Bench
67.7%自己申告
SWE-Bench ProPrinceton NLP (2024)
67.7%自己申告
QwenQoderBench
58.4%自己申告
DeepSWE 1.1
56.6%自己申告
NL2Repo
55.9%自己申告
Job Bench
53.4%自己申告
OneMillion Bench
52.5%自己申告
Agents' Last Exam
52.4%自己申告
MLS-Bench Lite
41.0%自己申告
AutomationBench
27.3%自己申告
Biology
GPQANYU + Cohere + Anthropic (2023)
92.6%自己申告
Code
Vision2Web
69.0%自己申告
Finance
PRBench-Finance
58.3%自己申告
General
MRCR v2 (8-needle)
92.9%自己申告
IFBench
82.8%自己申告
MMMU-Pro
82.3%自己申告
LongBench v2
66.3%自己申告
Grounding
ScreenSpot Pro
84.5%自己申告
Healthcare
HealthBench
60.2%自己申告
Knowledge
PLawBench
73.2%自己申告
PRBench-Legal
57.6%自己申告
Long Context
LVBench
81.8%自己申告
Math
Humanity's Last Exam (with tools, text-only)
56.2%自己申告
Humanity's Last Exam
43.6%自己申告
Multimodal
VideoMME w sub.
90.4%自己申告
PerceptionBench
63.5%自己申告
Reasoning
ERQA
77.8%自己申告
Spatial Reasoning
RealWorldQA
88.0%自己申告
AA評価指数
(Artificial Analysis)Coding Index(Artificial Analysis)71.8
Intelligence Index(Artificial Analysis)58.1
Gpqa(NYU + Cohere + Anthropic (2023))0.9
Terminalbench V2 10.8
Lcr(Artificial Analysis)0.7
Scicode(UIUC + Argonne National Lab (2024))0.5
Tau Banking0.5
Hle(Center for AI Safety + Scale AI (2025))0.4
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Physics90
Biology90
Chemistry90
Video90
Multimodal80
Search80
Spatial Reasoning80
Instruction Following80
Grounding80
Vision80
Long Context70
Reasoning70
Structured Output70
General70
Agents70
Code70
Productivity60
Healthcare60
Tool Calling60
Math50
価格設定
入力価格$2 / 1Mトークン
出力価格$6 / 1Mトークン
混合価格(3:1)$3 / 1Mトークン
キャッシュ読み取り価格$0.25 / 1Mトークン
キャッシュ書き込み価格$2.5 / 1Mトークン
速度
トークン/秒45.2
初トークン遅延1.95s
初回答遅延46.22s
プロバイダー価格ランキング
プロバイダー価格ランキング
13 プロバイダー
最安: AIHubMix最高: Charm Hyper
プロバイダー入力出力
1AIHubMix最安
$1.69
$5.07
2Alibaba (China)
$1.77744
$5.33231
3LLM Gateway
$1.815
$5.4461
4CrossModel
$1.88
$5.63
5Alibabaプライマリ
$2
$6
6NanoGPT
$2
$6
7OpenRouter
$2
$6
8OpenCode Go
$2
$6
9Kilo Gateway
$2
$6
10DigitalOcean
$2
$6
11Merge Gateway
$2
$6
12EmpirioLabs AI
$2
$6
13Charm Hyper
$2
$6
このモデルの異なるAPIプロバイダー間の価格を比較。