Qwen3.7 Max
AlibabaQwenProprietary
説明
Qwen3.7 Max is Alibaba Cloud Qwen Team's proprietary flagship model for agent-driven workflows. It is designed for coding agents, office automation, MCP and multi-agent orchestration, and long-horizon autonomous execution, with a 1 million token context window and up to 65,536 output tokens. Qwen reports strong agentic coding results including 69.7 on Terminal-Bench 2.0-Terminus, 80.4 on SWE-bench Verified, 60.6 on SWE-Pro, and 78.3 on SWE-Multilingual, alongside 92.4 on GPQA Diamond and 97.1 on HMMT 2026 Feb.
リリース日
2026-05-19
パラメータ
—
コンテキスト長
1.0M
モダリティ
text
能力レーダー
45
general
63
coding
92
reasoning
67
science
70
agents
0
multimodal
ランキング
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Agents
QwenSVG
1608.00 / 2000自己申告
Kernel Bench L3
96.0%自己申告
SpreadSheetBench-v1
87.0%自己申告
CoWorkBench
67.2%自己申告
MCP-Mark
60.8%自己申告
QwenWorldBench
57.3%自己申告
Finance Agent v2
48.4%
VITA-Bench
47.9%自己申告
Chat
IFEvalGoogle Research (2023)
94.3%自己申告
Code
QwenWebBench
1568.00 / 2000自己申告
Claw-Eval
65.2%自己申告
ZClawBench
64.3%自己申告
SkillsBench
59.2%自己申告
NL2Repo
47.2%自己申告
General
MAXIFE
89.2%自己申告
Include
86.2%自己申告
NOVA-63
59.0%自己申告
Instruction Following
IFBench
79.1%自己申告
Language
MMLU-Redux
95.0%自己申告
MMMLU
90.3%自己申告
MMLU-Pro
89.6%自己申告
MMLU-ProX
87.0%自己申告
WMT24++
85.8%自己申告
Long Context
MRCR 128K (8-needle)
90.4%自己申告
Math
HMMT Feb 26
97.1%自己申告
IMO-AnswerBench
90.0%自己申告
PolyMATH
86.5%自己申告
LiveBench
74.3%
MathArena Apex
44.5%自己申告
Reasoning
GPQANYU + Cohere + Anthropic (2023)
92.4%自己申告
LiveCodeBench v6
91.6%自己申告
Global PIQA
91.4%自己申告
SWE-Bench Verified
80.4%自己申告
SWE-bench Multilingual
78.3%自己申告
MCP Atlas
76.4%自己申告
SuperGPQA
73.6%自己申告
Terminal-Bench 2.0Stanford × Laude Institute (2026)
69.7%自己申告
SWE-Bench ProPrinceton NLP (2024)
60.6%自己申告
SciCode
53.5%自己申告
Humanity's Last Exam
41.4%自己申告
CritPT
11.4%自己申告
Tool Calling
BFCL-V4
75.0%自己申告
AA評価指数
(Artificial Analysis)Tau2(Sierra + U Toronto + Vector Institute (2025))94.7
Gpqa(NYU + Cohere + Anthropic (2023))92.3
Ifbench(Google Research (2023))80.5
Lcr(Artificial Analysis)74.7
Terminalbench V2 174.5
Coding Index(Artificial Analysis)66.0
Terminalbench Hard(Stanford × Laude Institute (2026))50.8
Scicode(UIUC + Argonne National Lab (2024))48.8
Intelligence Index(Artificial Analysis)46.7
Hle(Center for AI Safety + Scale AI (2025))40.5
Tau Banking11.8
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Chat90
Language90
Multimodal90
Spatial Reasoning90
Structured Output90
Instruction Following90
Legal80
Physics80
Productivity80
Frontend Development80
Healthcare80
Math70
Reasoning70
Finance70
General70
Biology70
Chemistry70
Code70
Economics70
Tool Calling70
Agents60
Vision60
価格設定
入力価格$2.5 / 1Mトークン
出力価格$7.5 / 1Mトークン
混合価格(3:1)$3.75 / 1Mトークン
キャッシュ読み取り価格$0.5 / 1Mトークン
キャッシュ書き込み価格$3.125 / 1Mトークン
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
24 プロバイダー
最安: Novita最高: Modelis
プロバイダー入力出力
1Novita最安
$0
$0
2Together
$0
$0.00001
3Merge Gateway
$0.825
$2.4755
4NovitaAI
$1.25
$3.75
5Kilo Gateway
$1.25
$3.75
6DevPass (LLM Gateway)
$1.25
$3.75
7OrcaRouter
$1.25
$3.75
8Pioneer
$1.25
$3.75
9OpenRouter
$1.475
$4.425
10CrossModel
$1.504
$4.504
11AIHubMix
$1.69
$5.07
12Alibabaプライマリ
$2.5
$7.5
13NanoGPT
$2.5
$7.5
14Abacus
$2.5
$7.5
15OpenCode Go
$2.5
$7.5
16Alibaba (China)
$2.5
$7.5
17ZenMux
$2.5
$7.5
18Alibaba Coding Plan
$2.5
$7.5
19Requesty
$2.5
$7.5
20Alibaba Coding Plan (China)
$2.5
$7.5
21EmpirioLabs AI
$2.5
$7.5
22Charm Hyper
$2.5
$7.5
23Impossibl
$2.5
$7.5
24Modelis
$3
$9
このモデルの異なるAPIプロバイダー間の価格を比較。