Qwen3.5 122B A10B (Non-reasoning)
AlibabaQwenOpen WeightApache 2.0 · Uso Comercial
Descripción
Qwen3.5-122B-A10B is a multimodal Mixture-of-Experts model with 122 billion total parameters and 10 billion activated parameters. It combines strong reasoning, coding, long-context, and visual understanding performance with production-friendly efficiency and a native 262K context window.
Fecha de lanzamiento
2026-02-24
Parámetros
122.0B
Longitud del contexto
262K
Modalidades
audio, image, text, video
Radar de capacidades
17
general
43
coding
83
reasoning
62
science
60
agents
80
multimodal
Rankings
| Dominio | #Posición | Puntuación | Fuente |
|---|---|---|---|
| Capacidad agéntica | 47 | 49.0 | LS |
| Ranking de codificación | 263 | 55.0 | AA |
| Ranking general | 245 | 48.0 | AA |
| Ranking multimodal | 43 | 59.0 | LS |
| Ciencia | 226 | 55.0 | AA |
Puntuaciones de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
79.5%Aut.
VITA-Bench
33.6%Aut.
DeepPlanning
24.1%Aut.
Chat
IFEvalGoogle Research (2023)
93.4%Aut.
Multi-Challenge
61.5%Aut.
Code
FullStackBench en
62.6%Aut.
FullStackBench zh
58.7%Aut.
General
C-Eval
91.9%Aut.
MAXIFE
87.9%Aut.
Include
82.8%Aut.
AndroidWorld_SR
66.4%Aut.
NOVA-63
58.6%Aut.
Healthcare
MedXpertQA
67.3%Aut.
PMC-VQA
63.3%Aut.
Instruction Following
IFBench
76.1%Aut.
Language
MMLU-Redux
94.0%Aut.
MMMLU
86.7%Aut.
MMLU-Pro
86.7%Aut.
MMLU-ProX
82.2%Aut.
WMT24++
78.3%Aut.
Long Context
LongBench v2
60.2%Aut.
Math
HMMT 2025
91.4%Aut.
HMMT25
90.3%Aut.
MathVista-Mini
87.4%Aut.
MathVision
86.2%Aut.
DynaMath
85.9%Aut.
CodeForces
0.85 / 3000Aut.
PolyMATH
68.9%Aut.
Multimodal
VideoMME w/o sub.
83.9%Aut.
MMMU
83.9%Aut.
VideoMMMU
82.0%Aut.
OSWorld-Verified
58.0%Aut.
TIR-Bench
53.2%Aut.
Reasoning
Global PIQA
88.4%Aut.
GPQANYU + Cohere + Anthropic (2023)
86.6%Aut.
LiveCodeBench v6
78.9%Aut.
CharXiv-R
77.2%Aut.
SWE-Bench Verified
72.0%Aut.
BrowseComp-zh
69.9%Aut.
SuperGPQA
67.1%Aut.
AA-LCR
66.9%Aut.
BrowseCompOpenAI (2025)
63.8%Aut.
Terminal-Bench 2.0Stanford × Laude Institute (2026)
49.4%Aut.
Humanity's Last Exam
47.5%Aut.
Seal-0
44.1%Aut.
OJBench
39.5%Aut.
Search
WideSearch
60.5%Aut.
Tool Calling
BFCL-V4
72.2%Aut.
Video
MLVU
87.3%Aut.
Vision
CountBench
0.97 / 100Aut.
VLMsAreBlind
96.7%Aut.
AI2D
93.3%Aut.
V*
93.2%Aut.
MMBench-V1.1
92.8%Aut.
OCRBench
92.1%Aut.
RefCOCO-avg
0.91 / 100Aut.
OmniDocBench 1.5
89.8%Aut.
VideoMME w sub.
87.3%Aut.
RealWorldQA
85.1%Aut.
EmbSpatialBench
0.84 / 100Aut.
MMStar
82.9%Aut.
CC-OCR
81.8%Aut.
SlakeVQA
81.6%Aut.
LingoQA
80.8%Aut.
MMMU-Pro
76.9%Aut.
MVBench
76.6%Aut.
MMVU
74.7%Aut.
LVBench
74.4%Aut.
ScreenSpot Pro
70.4%Aut.
RefSpatialBench
0.69 / 100Aut.
Hallusion Bench
67.6%Aut.
ERQA
62.0%Aut.
SimpleVQA
0.62 / 100Aut.
MMLongBench-Doc
0.59 / 100Aut.
ODinW
44.5%Aut.
BabyVision
40.2%Aut.
SUNRGBD
0.36 / 100Aut.
ZEROBench-Sub
0.36 / 100Aut.
Nuscene
15.4%Aut.
Hypersim
0.13 / 100Aut.
ZEROBench
0.09 / 100Aut.
Índices de evaluación AA
(Artificial Analysis)Tau2(Sierra + U Toronto + Vector Institute (2025))84.5
Gpqa(NYU + Cohere + Anthropic (2023))82.7
Lcr(Artificial Analysis)61.3
Ifbench(Google Research (2023))50.8
Terminalbench V2 147.2
Coding Index(Artificial Analysis)43.3
Terminalbench Hard(Stanford × Laude Institute (2026))29.5
Intelligence Index(Artificial Analysis)17.7
Hle(Center for AI Safety + Scale AI (2025))15.9
Tau Banking10.3
Puntuaciones por categoría LLM Stats
(LLM Stats (zeroeval))Biology90
Chat80
Image To Text80
Instruction Following80
Language80
Legal80
Math80
Physics80
Structured Output80
Embodied80
Finance80
Grounding80
Healthcare80
Chemistry80
Text-to-image80
Video80
Long Context70
Multimodal70
Reasoning70
Spatial Reasoning70
Frontend Development70
General70
Economics70
Vision70
Search60
Agents60
Code60
Communication60
Tool Calling60
Spatial20
3d20
Precios
Precio de entrada$0.4 / 1M tokens
Precio de salida$3.2 / 1M tokens
Precio mixto (3:1)$1.1 / 1M tokens
Velocidad
Tokens/seg146.1
Retraso del primer token1.00s
Tiempo hasta la respuesta1.00s
Ranking de Precios por Proveedor
Ranking de Precios por Proveedor
2 proveedores
Más barato: DeepInfraMás caro: Alibaba
ProveedorEntradaSalida
1DeepInfraMás barato
$0
$0
2AlibabaPRINCIPAL
$0.4
$3.2
Comparar precios entre diferentes proveedores de API para este modelo.