Qwen3.8 Max
AlibabaQwenOpen WeightQwen3.8-Max License · Uso Comercial
Descripción
Qwen3.8 Max, released as the Qwen3.8-2.4T-A95B open-weight checkpoint, is Qwen's flagship mixture-of-experts model for coding, research, professional workflows, and long-horizon agents. It has 2.4 trillion total parameters with 95 billion active parameters, a native 262,144-token context window extensible to about 1 million tokens, and always-on thinking. QwenCloud serves the same model in the Qwen3.8 Max API family with text and image input.
Fecha de lanzamiento
2026-08-03
Parámetros
2.4T
Longitud del contexto
1.0M
Modalidades
image, pdf, text, video
Radar de capacidades
55
general
69
coding
93
reasoning
69
science
60
agents
90
multimodal
Rankings
| Dominio | #Posición | Puntuación | Fuente |
|---|---|---|---|
| Capacidad agéntica | 4 | 79.0 | LS |
| Ranking de codificación | 26 | 91.0 | AA |
| Ranking general | 8 | 92.0 | AA |
| Razonamiento matemático | 17 | 91.0 | LB |
| Ranking multimodal | 10 | 65.0 | LS |
| Razonamiento | 13 | 88.0 | LB |
| Ciencia | 26 | 87.0 | AA |
Puntuaciones de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Agents
QwenReactBench
1724.00 / 2000Aut.
QwenSVG
1713.00 / 2000Aut.
PaperBench
93.0%Aut.
Terminal-Bench 2.1
86.6%Aut.
OSWorld-Verified
86.1%Aut.
AndroidWorld
85.3%Aut.
WideSearch
81.9%Aut.
QwenSWEBench
80.7%Aut.
MobileWorld
77.8%Aut.
AndroidBench
75.1%Aut.
CoWorkBench
74.8%Aut.
FrontierSWE
73.5%Aut.
Toolathlon
72.5%Aut.
SkillsBench
70.2%Aut.
Workspace Bench
67.7%Aut.
SWE-Bench ProPrinceton NLP (2024)
67.7%Aut.
QwenQoderBench
58.4%Aut.
DeepSWE 1.1
56.6%Aut.
NL2Repo
55.9%Aut.
Job Bench
53.4%Aut.
OneMillion Bench
52.5%Aut.
Agents' Last Exam
52.4%Aut.
MLS-Bench Lite
41.0%Aut.
AutomationBench
27.3%Aut.
Biology
GPQANYU + Cohere + Anthropic (2023)
92.6%Aut.
Code
Vision2Web
69.0%Aut.
Finance
PRBench-Finance
58.3%Aut.
General
MRCR v2 (8-needle)
92.9%Aut.
IFBench
82.8%Aut.
MMMU-Pro
82.3%Aut.
LongBench v2
66.3%Aut.
Grounding
ScreenSpot Pro
84.5%Aut.
Healthcare
HealthBench
60.2%Aut.
Knowledge
PLawBench
73.2%Aut.
PRBench-Legal
57.6%Aut.
Long Context
LVBench
81.8%Aut.
Math
Humanity's Last Exam (with tools, text-only)
56.2%Aut.
Humanity's Last Exam
43.6%Aut.
Multimodal
VideoMME w sub.
90.4%Aut.
PerceptionBench
63.5%Aut.
Reasoning
ERQA
77.8%Aut.
Spatial Reasoning
RealWorldQA
88.0%Aut.
Índices de evaluación AA
(Artificial Analysis)Coding Index(Artificial Analysis)71.8
Intelligence Index(Artificial Analysis)58.1
Gpqa(NYU + Cohere + Anthropic (2023))0.9
Terminalbench V2 10.8
Lcr(Artificial Analysis)0.7
Scicode(UIUC + Argonne National Lab (2024))0.5
Tau Banking0.5
Hle(Center for AI Safety + Scale AI (2025))0.4
Puntuaciones por categoría LLM Stats
(LLM Stats (zeroeval))Physics90
Biology90
Chemistry90
Video90
Multimodal80
Search80
Spatial Reasoning80
Instruction Following80
Grounding80
Vision80
Long Context70
Reasoning70
Structured Output70
General70
Agents70
Code70
Productivity60
Healthcare60
Tool Calling60
Math50
Precios
Precio de entrada$2 / 1M tokens
Precio de salida$6 / 1M tokens
Precio mixto (3:1)$3 / 1M tokens
Precio de lectura caché$0.25 / 1M tokens
Precio de escritura caché$2.5 / 1M tokens
Velocidad
Tokens/seg45.2
Retraso del primer token1.95s
Tiempo hasta la respuesta46.22s
Ranking de Precios por Proveedor
Ranking de Precios por Proveedor
13 proveedores
Más barato: AIHubMixMás caro: Charm Hyper
ProveedorEntradaSalida
1AIHubMixMás barato
$1.69
$5.07
2Alibaba (China)
$1.77744
$5.33231
3LLM Gateway
$1.815
$5.4461
4CrossModel
$1.88
$5.63
5AlibabaPRINCIPAL
$2
$6
6NanoGPT
$2
$6
7OpenRouter
$2
$6
8OpenCode Go
$2
$6
9Kilo Gateway
$2
$6
10DigitalOcean
$2
$6
11Merge Gateway
$2
$6
12EmpirioLabs AI
$2
$6
13Charm Hyper
$2
$6
Comparar precios entre diferentes proveedores de API para este modelo.