Qwen3.5 397B A17B (Non-reasoning)
AlibabaQwenOpen WeightApache 2.0 · Uso Comercial
Descripción
Qwen3.5-397B-A17B is Qwen's flagship Mixture-of-Experts model with 397 billion total parameters and 17 billion activated parameters. It delivers state-of-the-art performance across knowledge, reasoning, coding, mathematics, multilingual understanding, instruction following, long context, and agent tasks.
Fecha de lanzamiento
2026-02-16
Parámetros
397.0B
Longitud del contexto
262K
Modalidades
audio, image, text, video
Radar de capacidades
21
general
70
coding
86
reasoning
66
science
60
agents
70
multimodal
Rankings
| Dominio | #Posición | Puntuación | Fuente |
|---|---|---|---|
| Capacidad agéntica | 17 | 62.0 | LS |
| Ranking de codificación | 216 | 64.0 | AA |
| Ranking general | 211 | 52.0 | AA |
| Ciencia | 195 | 60.0 | AA |
Puntuaciones de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
86.7%Aut.
VITA-Bench
49.7%Aut.
MCP-Mark
46.1%Aut.
Toolathlon
38.3%Aut.
DeepPlanning
34.3%Aut.
Chat
IFEvalGoogle Research (2023)
92.6%Aut.
Multi-Challenge
67.6%Aut.
Code
SecCodeBench
68.3%Aut.
General
C-Eval
93.0%Aut.
MAXIFE
88.2%Aut.
Include
85.6%Aut.
NOVA-63
59.1%Aut.
Instruction Following
IFBench
76.5%Aut.
Language
MMLU-Redux
94.9%Aut.
MMMLU
88.5%Aut.
MMLU-Pro
87.8%Aut.
MMLU-ProX
84.7%Aut.
WMT24++
78.9%Aut.
Long Context
LongBench v2
63.2%Aut.
Math
HMMT 2025
94.8%Aut.
HMMT25
92.7%Aut.
AIME 2026
91.3%Aut.
IMO-AnswerBench
80.9%Aut.
PolyMATH
73.3%Aut.
Reasoning
Global PIQA
89.8%Aut.
GPQANYU + Cohere + Anthropic (2023)
88.4%Aut.
LiveCodeBench v6
83.6%Aut.
SWE-Bench Verified
76.4%Aut.
SuperGPQA
70.4%Aut.
BrowseComp-zh
70.3%Aut.
SWE-bench Multilingual
69.3%Aut.
BrowseCompOpenAI (2025)
69.0%Aut.
AA-LCR
68.7%Aut.
Terminal-Bench 2.0Stanford × Laude Institute (2026)
52.5%Aut.
Seal-0
46.9%Aut.
Humanity's Last Exam
28.7%Aut.
Search
WideSearch
74.0%Aut.
Tool Calling
BFCL-V4
72.9%Aut.
Índices de evaluación AA
(Artificial Analysis)Gpqa(NYU + Cohere + Anthropic (2023))86.1
Tau2(Sierra + U Toronto + Vector Institute (2025))83.9
Lcr(Artificial Analysis)64.3
Ifbench(Google Research (2023))51.6
Terminalbench Hard(Stanford × Laude Institute (2026))35.6
Intelligence Index(Artificial Analysis)21.4
Hle(Center for AI Safety + Scale AI (2025))19.8
Puntuaciones por categoría LLM Stats
(LLM Stats (zeroeval))Language90
Biology90
Chat80
Instruction Following80
Legal80
Math80
Physics80
Structured Output80
Finance80
Frontend Development80
Healthcare80
Chemistry80
Long Context70
Multimodal70
Reasoning70
Search70
Spatial Reasoning70
General70
Code70
Communication70
Economics70
Agents60
Tool Calling60
Vision50
Precios
Precio de entrada$0.6 / 1M tokens
Precio de salida$3.6 / 1M tokens
Precio mixto (3:1)$1.35 / 1M tokens
Velocidad
Tokens/seg86.1
Retraso del primer token1.70s
Tiempo hasta la respuesta1.70s
Ranking de Precios por Proveedor
Ranking de Precios por Proveedor
2 proveedores
Más barato: DeepInfraMás caro: Alibaba
ProveedorEntradaSalida
1DeepInfraMás barato
$0
$0
2AlibabaPRINCIPAL
$0.6
$3.6
Comparar precios entre diferentes proveedores de API para este modelo.