Qwen3.5 122B A10B (Non-reasoning)
AlibabaQwen오픈 웨이트Apache 2.0 · 상업적 사용 가능
설명
Qwen3.5-122B-A10B is a multimodal Mixture-of-Experts model with 122 billion total parameters and 10 billion activated parameters. It combines strong reasoning, coding, long-context, and visual understanding performance with production-friendly efficiency and a native 262K context window.
출시일
2026-02-24
파라미터
122.0B
컨텍스트 길이
262K
모달리티
audio, image, text, video
능력 레이더
17
general
43
coding
83
reasoning
62
science
60
agents
80
multimodal
랭킹
벤치마크 점수 (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
79.5%자체 보고
VITA-Bench
33.6%자체 보고
DeepPlanning
24.1%자체 보고
Chat
IFEvalGoogle Research (2023)
93.4%자체 보고
Multi-Challenge
61.5%자체 보고
Code
FullStackBench en
62.6%자체 보고
FullStackBench zh
58.7%자체 보고
General
C-Eval
91.9%자체 보고
MAXIFE
87.9%자체 보고
Include
82.8%자체 보고
AndroidWorld_SR
66.4%자체 보고
NOVA-63
58.6%자체 보고
Healthcare
MedXpertQA
67.3%자체 보고
PMC-VQA
63.3%자체 보고
Instruction Following
IFBench
76.1%자체 보고
Language
MMLU-Redux
94.0%자체 보고
MMMLU
86.7%자체 보고
MMLU-Pro
86.7%자체 보고
MMLU-ProX
82.2%자체 보고
WMT24++
78.3%자체 보고
Long Context
LongBench v2
60.2%자체 보고
Math
HMMT 2025
91.4%자체 보고
HMMT25
90.3%자체 보고
MathVista-Mini
87.4%자체 보고
MathVision
86.2%자체 보고
DynaMath
85.9%자체 보고
CodeForces
0.85 / 3000자체 보고
PolyMATH
68.9%자체 보고
Multimodal
VideoMME w/o sub.
83.9%자체 보고
MMMU
83.9%자체 보고
VideoMMMU
82.0%자체 보고
OSWorld-Verified
58.0%자체 보고
TIR-Bench
53.2%자체 보고
Reasoning
Global PIQA
88.4%자체 보고
GPQANYU + Cohere + Anthropic (2023)
86.6%자체 보고
LiveCodeBench v6
78.9%자체 보고
CharXiv-R
77.2%자체 보고
SWE-Bench Verified
72.0%자체 보고
BrowseComp-zh
69.9%자체 보고
SuperGPQA
67.1%자체 보고
AA-LCR
66.9%자체 보고
BrowseCompOpenAI (2025)
63.8%자체 보고
Terminal-Bench 2.0Stanford × Laude Institute (2026)
49.4%자체 보고
Humanity's Last Exam
47.5%자체 보고
Seal-0
44.1%자체 보고
OJBench
39.5%자체 보고
Search
WideSearch
60.5%자체 보고
Tool Calling
BFCL-V4
72.2%자체 보고
Video
MLVU
87.3%자체 보고
Vision
CountBench
0.97 / 100자체 보고
VLMsAreBlind
96.7%자체 보고
AI2D
93.3%자체 보고
V*
93.2%자체 보고
MMBench-V1.1
92.8%자체 보고
OCRBench
92.1%자체 보고
RefCOCO-avg
0.91 / 100자체 보고
OmniDocBench 1.5
89.8%자체 보고
VideoMME w sub.
87.3%자체 보고
RealWorldQA
85.1%자체 보고
EmbSpatialBench
0.84 / 100자체 보고
MMStar
82.9%자체 보고
CC-OCR
81.8%자체 보고
SlakeVQA
81.6%자체 보고
LingoQA
80.8%자체 보고
MMMU-Pro
76.9%자체 보고
MVBench
76.6%자체 보고
MMVU
74.7%자체 보고
LVBench
74.4%자체 보고
ScreenSpot Pro
70.4%자체 보고
RefSpatialBench
0.69 / 100자체 보고
Hallusion Bench
67.6%자체 보고
ERQA
62.0%자체 보고
SimpleVQA
0.62 / 100자체 보고
MMLongBench-Doc
0.59 / 100자체 보고
ODinW
44.5%자체 보고
BabyVision
40.2%자체 보고
SUNRGBD
0.36 / 100자체 보고
ZEROBench-Sub
0.36 / 100자체 보고
Nuscene
15.4%자체 보고
Hypersim
0.13 / 100자체 보고
ZEROBench
0.09 / 100자체 보고
AA 평가 지수
(Artificial Analysis)Tau2(Sierra + U Toronto + Vector Institute (2025))84.5
Gpqa(NYU + Cohere + Anthropic (2023))82.7
Lcr(Artificial Analysis)61.3
Ifbench(Google Research (2023))50.8
Terminalbench V2 147.2
Coding Index(Artificial Analysis)43.3
Terminalbench Hard(Stanford × Laude Institute (2026))29.5
Intelligence Index(Artificial Analysis)17.7
Hle(Center for AI Safety + Scale AI (2025))15.9
Tau Banking10.3
LLM Stats 카테고리 점수
(LLM Stats (zeroeval))Biology90
Chat80
Image To Text80
Instruction Following80
Language80
Legal80
Math80
Physics80
Structured Output80
Embodied80
Finance80
Grounding80
Healthcare80
Chemistry80
Text-to-image80
Video80
Long Context70
Multimodal70
Reasoning70
Spatial Reasoning70
Frontend Development70
General70
Economics70
Vision70
Search60
Agents60
Code60
Communication60
Tool Calling60
Spatial20
3d20
가격
입력 가격$0.4 / 1M 토큰
출력 가격$3.2 / 1M 토큰
혼합 가격 (3:1)$1.1 / 1M 토큰
속도
토큰/초147.0
첫 토큰 지연1.00s
첫 응답 지연1.00s
공급자 가격 순위
공급자 가격 순위
2개 공급자
최저가: DeepInfra최고가: Alibaba
공급자입력출력
1DeepInfra최저가
$0
$0
2Alibaba주요
$0.4
$3.2
이 모델의 다양한 API 공급자 간 가격 비교.