Qwen3.6 Plus
AlibabaQwenProprietary
Descripción
Qwen3.6 Plus is Alibaba's next-generation flagship model featuring a 1 million token native context window, up to 65,536 output tokens, and always-on chain-of-thought reasoning. It uses a next-generation hybrid architecture optimized for efficiency and scalability. It leads on Terminal-Bench 2.0 agentic coding (61.6), surpassing Claude 4.5 Opus, and achieves strong results on document understanding (OmniDocBench 91.2) and multimodal reasoning (MMMU 86.0). Compared to Qwen 3.5, it is significantly more decisive in reasoning, using fewer tokens on straightforward tasks with better agent stability.
Fecha de lanzamiento
2026-04-02
Parámetros
—
Longitud del contexto
1.0M
Modalidades
image, text, video
Radar de capacidades
38
general
52
coding
88
reasoning
59
science
60
agents
90
multimodal
Rankings
| Dominio | #Posición | Puntuación | Fuente |
|---|---|---|---|
| Capacidad agéntica | 97 | 40.0 | LS |
| Ranking de codificación | 89 | 74.0 | AA |
| Ranking general | 55 | 79.0 | AA |
| Razonamiento matemático | 29 | 84.0 | LB |
| Ranking multimodal | 2 | 75.0 | LS |
| Razonamiento | 29 | 76.0 | LB |
| Ciencia | 103 | 70.0 | AA |
Puntuaciones de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Agents
WideSearch
74.3%Aut.
MCP Atlas
74.1%Aut.
TAU3-Bench
70.7%Aut.
OSWorld-Verified
62.5%Aut.
TIR-Bench
61.6%Aut.
Terminal-Bench 2.0Stanford × Laude Institute (2026)
61.6%Aut.
Claw-Eval
58.7%Aut.
SWE-Bench ProPrinceton NLP (2024)
56.6%Aut.
MCP-Mark
48.2%Aut.
SkillsBench
45.7%Aut.
VITA-Bench
44.3%Aut.
DeepPlanning
41.5%Aut.
Finance Agent v2
40.8%
Toolathlon
39.8%Aut.
NL2Repo
37.9%Aut.
FrontierSWE
22.0%
Biology
GPQANYU + Cohere + Anthropic (2023)
90.4%Aut.
Chemistry
SuperGPQA
71.6%Aut.
Code
SWE-Bench Verified
78.8%Aut.
SWE-bench Multilingual
73.8%Aut.
Finance
MMLU-Pro
88.5%Aut.
MMLU-ProX
84.7%Aut.
General
MMLU-Redux
94.5%Aut.
IFEvalGoogle Research (2023)
94.3%Aut.
C-Eval
93.3%Aut.
Global PIQA
89.8%Aut.
MMMLU
89.5%Aut.
MAXIFE
88.2%Aut.
LiveCodeBench v6
87.1%Aut.
MMMU
86.0%Aut.
Include
85.1%Aut.
MMStar
83.3%Aut.
MMMU-Pro
78.8%Aut.
IFBench
74.2%Aut.
LiveBench
70.9%
SimpleVQA
0.67 / 100Aut.
LongBench v2
62.0%Aut.
NOVA-63
57.9%Aut.
Grounding
RefCOCO-avg
0.94 / 100Aut.
ScreenSpot Pro
68.2%Aut.
Healthcare
VideoMMMU
84.0%Aut.
Language
WMT24++
84.3%Aut.
Long Context
MLVU
86.7%Aut.
AA-LCR
68.3%Aut.
MMLongBench-Doc
0.62 / 100Aut.
Math
HMMT 2025
96.7%Aut.
AIME 2026
95.3%Aut.
HMMT25
94.6%Aut.
We-Math
89.0%Aut.
DynaMath
88.0%Aut.
MathVision
88.0%Aut.
HMMT Feb 26
87.8%Aut.
IMO-AnswerBench
83.8%Aut.
PolyMATH
77.4%Aut.
Humanity's Last Exam
28.8%Aut.
Multimodal
V*
96.9%Aut.
AI2D
94.4%Aut.
OmniDocBench 1.5
91.2%Aut.
Video-MME
84.2%Aut.
CC-OCR
83.4%Aut.
CharXiv-R
81.5%Aut.
Reasoning
CountBench
0.98 / 100Aut.
ERQA
65.7%Aut.
Spatial Reasoning
RealWorldQA
85.4%Aut.
Vision
ODinW
51.8%Aut.
Índices de evaluación AA
(Artificial Analysis)Coding Index(Artificial Analysis)54.5
Intelligence Index(Artificial Analysis)40.5
Tau2(Sierra + U Toronto + Vector Institute (2025))1.0
Gpqa(NYU + Cohere + Anthropic (2023))0.9
Ifbench(Google Research (2023))0.8
Lcr(Artificial Analysis)0.7
Terminalbench V2 10.6
Terminalbench Hard(Stanford × Laude Institute (2026))0.4
Scicode(UIUC + Argonne National Lab (2024))0.4
Hle(Center for AI Safety + Scale AI (2025))0.3
Tau Banking0.2
Puntuaciones por categoría LLM Stats
(LLM Stats (zeroeval))Language90
Biology90
Video90
Legal80
Math80
Multimodal80
Physics80
Reasoning80
Spatial Reasoning80
Structured Output80
Instruction Following80
Frontend Development80
Grounding80
Healthcare80
Chemistry80
Text-to-image80
Vision80
Long Context70
Search70
Image To Text70
Finance70
General70
Economics70
Code60
Tool Calling60
Agents50
Precios
Precio de entrada$0.5 / 1M tokens
Precio de salida$3 / 1M tokens
Precio mixto (3:1)$1.125 / 1M tokens
Precio de lectura caché$0.05 / 1M tokens
Precio de escritura caché$0.625 / 1M tokens
Velocidad
Tokens/seg0.0
Retraso del primer token0.00s
Tiempo hasta la respuesta0.00s
Ranking de Precios por Proveedor
Ranking de Precios por Proveedor
17 proveedores
Más barato: TogetherMás caro: Charm Hyper
ProveedorEntradaSalida
1TogetherMás barato
$0
$0
2Novita
$0
$0
3Merge Gateway
$0.276
$1.651
4AIHubMix
$0.28
$1.69
5CrossModel
$0.32
$1.88
6OpenRouter
$0.325
$1.95
7Kilo Gateway
$0.325
$1.95
8Pioneer
$0.325
$1.95
9AlibabaPRINCIPAL
$0.5
$3
10OpenCode Go
$0.5
$3
11Alibaba (China)
$0.5
$3
12ZenMux
$0.5
$3
13OpenCode Zen
$0.5
$3
14LLM Gateway
$0.5
$3
15OrcaRouter
$0.5
$3
16EmpirioLabs AI
$0.5
$3
17Charm Hyper
$2
$6
Comparar precios entre diferentes proveedores de API para este modelo.