メインコンテンツへスキップ

Qwen3.8 Max (0902)

AlibabaQwenオープンウエイトQwen3.8-Max License · 商用利用可

説明

Qwen3.8 Max, released as the Qwen3.8-2.4T-A95B open-weight checkpoint, is Qwen's flagship mixture-of-experts model for coding, research, professional workflows, and long-horizon agents. It has 2.4 trillion total parameters with 95 billion active parameters, a native 262,144-token context window extensible to about 1 million tokens, and always-on thinking. QwenCloud serves the same model in the Qwen3.8 Max API family with text and image input.

リリース日
2026-09-02
パラメータ
2.4T
コンテキスト長
1.0M
モダリティ
image, pdf, text, video

能力レーダー

45
general
72
coding
93
reasoning
69
science
60
agents
90
multimodal

ランキング

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Agents

QwenSVG1713.00 / 2000自己申告
AndroidWorld85.3%自己申告
MobileWorld77.8%自己申告
AndroidBench75.1%自己申告
CoWorkBench74.8%自己申告
Toolathlon72.5%自己申告
Workspace Bench67.7%自己申告
Job Bench53.4%自己申告
OneMillion Bench52.5%自己申告
Agents' Last Exam52.4%自己申告
MLS-Bench Lite41.0%自己申告
AutomationBench27.3%自己申告

Code

QwenReactBench1724.00 / 2000自己申告
PaperBench93.0%自己申告
QwenSWEBench80.7%自己申告
FrontierSWE73.5%自己申告
SkillsBench70.2%自己申告
Vision2Web69.0%自己申告
QwenQoderBench58.4%自己申告
DeepSWE 1.156.6%自己申告
NL2Repo55.9%自己申告

Healthcare

HealthBench60.2%自己申告

Instruction Following

IFBench82.8%自己申告

Knowledge

PLawBench73.2%自己申告
PRBench-Finance58.3%自己申告
PRBench-Legal57.6%自己申告

Long Context

MRCR v2 (8-needle)92.9%自己申告
LongBench v266.3%自己申告

Multimodal

OSWorld-Verified86.1%自己申告

Reasoning

GPQANYU + Cohere + Anthropic (2023)92.6%自己申告
Terminal-Bench 2.186.6%自己申告
SWE-Bench ProPrinceton NLP (2024)67.7%自己申告
Humanity's Last Exam (with tools, text-only)56.2%自己申告
Humanity's Last Exam43.6%自己申告

Search

WideSearch81.9%自己申告

Vision

VideoMME w sub.90.4%自己申告
RealWorldQA88.0%自己申告
ScreenSpot Pro84.5%自己申告
MMMU-Pro82.3%自己申告
LVBench81.8%自己申告
ERQA77.8%自己申告
PerceptionBench63.5%自己申告

AA評価指数

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
92.8
Terminalbench V2 1
88.8
Lcr(Artificial Analysis)
80.3
Coding Index(Artificial Analysis)
76.2
Scicode(UIUC + Argonne National Lab (2024))
52.1
Tau Banking
47.8
Intelligence Index(Artificial Analysis)
45.4
Hle(Center for AI Safety + Scale AI (2025))
43.1
Terminalbench V4 0
38.9

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Physics
90
Biology
90
Chemistry
90
Video
90
Instruction Following
80
Multimodal
80
Search
80
Spatial Reasoning
80
Grounding
80
Vision
80
Long Context
70
Reasoning
70
Structured Output
70
General
70
Agents
70
Code
70
Productivity
60
Healthcare
60
Tool Calling
60
Math
50

価格設定

入力価格$2 / 1Mトークン
出力価格$6 / 1Mトークン
混合価格(3:1)$3 / 1Mトークン
キャッシュ読み取り価格$0.25 / 1Mトークン
キャッシュ書き込み価格$2.5 / 1Mトークン

速度

トークン/秒39.8
初トークン遅延1.81s
初回答遅延52.03s

プロバイダー価格ランキング

プロバイダー価格ランキング

6 プロバイダー

最安: DeepInfra最高: EmpirioLabs AI
プロバイダー入力出力
1DeepInfra最安
$0
$0
2Novita
$0
$0.00001
3Fireworks
$0
$0.00001
4Together
$0
$0.00001
5Alibabaプライマリ
$2
$6
6EmpirioLabs AI
$2
$6

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク