메인 콘텐츠로 건너뛰기

Qwen3.5 397B A17B (Non-reasoning)

AlibabaQwen오픈 웨이트Apache 2.0 · 상업적 사용 가능

설명

Qwen3.5-397B-A17B is Qwen's flagship Mixture-of-Experts model with 397 billion total parameters and 17 billion activated parameters. It delivers state-of-the-art performance across knowledge, reasoning, coding, mathematics, multilingual understanding, instruction following, long context, and agent tasks.

출시일
2026-02-16
파라미터
397.0B
컨텍스트 길이
262K
모달리티
audio, image, text, video

능력 레이더

21
general
70
coding
86
reasoning
66
science
60
agents
70
multimodal

랭킹

도메인#순위점수소스
에이전트형 역량18
62.0
LS
코딩 랭킹210
64.0
AA
종합 랭킹207
52.0
AA
과학189
60.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Agents

t2-bench86.7%자체 보고
VITA-Bench49.7%자체 보고
MCP-Mark46.1%자체 보고
Toolathlon38.3%자체 보고
DeepPlanning34.3%자체 보고

Chat

IFEvalGoogle Research (2023)92.6%자체 보고
Multi-Challenge67.6%자체 보고

Code

SecCodeBench68.3%자체 보고

General

C-Eval93.0%자체 보고
MAXIFE88.2%자체 보고
Include85.6%자체 보고
NOVA-6359.1%자체 보고

Instruction Following

IFBench76.5%자체 보고

Language

MMLU-Redux94.9%자체 보고
MMMLU88.5%자체 보고
MMLU-Pro87.8%자체 보고
MMLU-ProX84.7%자체 보고
WMT24++78.9%자체 보고

Long Context

LongBench v263.2%자체 보고

Math

HMMT 202594.8%자체 보고
HMMT2592.7%자체 보고
AIME 202691.3%자체 보고
IMO-AnswerBench80.9%자체 보고
PolyMATH73.3%자체 보고

Reasoning

Global PIQA89.8%자체 보고
GPQANYU + Cohere + Anthropic (2023)88.4%자체 보고
LiveCodeBench v683.6%자체 보고
SWE-Bench Verified76.4%자체 보고
SuperGPQA70.4%자체 보고
BrowseComp-zh70.3%자체 보고
SWE-bench Multilingual69.3%자체 보고
BrowseCompOpenAI (2025)69.0%자체 보고
AA-LCR68.7%자체 보고
Terminal-Bench 2.0Stanford × Laude Institute (2026)52.5%자체 보고
Seal-046.9%자체 보고
Humanity's Last Exam28.7%자체 보고

Search

WideSearch74.0%자체 보고

Tool Calling

BFCL-V472.9%자체 보고

AA 평가 지수

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
86.1
Tau2(Sierra + U Toronto + Vector Institute (2025))
83.9
Lcr(Artificial Analysis)
64.3
Ifbench(Google Research (2023))
51.6
Terminalbench Hard(Stanford × Laude Institute (2026))
35.6
Intelligence Index(Artificial Analysis)
21.4
Hle(Center for AI Safety + Scale AI (2025))
19.8

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Language
90
Biology
90
Chat
80
Instruction Following
80
Legal
80
Math
80
Physics
80
Structured Output
80
Finance
80
Frontend Development
80
Healthcare
80
Chemistry
80
Long Context
70
Multimodal
70
Reasoning
70
Search
70
Spatial Reasoning
70
General
70
Code
70
Communication
70
Economics
70
Agents
60
Tool Calling
60
Vision
50

가격

입력 가격$0.6 / 1M 토큰
출력 가격$3.6 / 1M 토큰
혼합 가격 (3:1)$1.35 / 1M 토큰

속도

토큰/초86.3
첫 토큰 지연1.69s
첫 응답 지연1.69s

공급자 가격 순위

공급자 가격 순위

2개 공급자

최저가: DeepInfra최고가: Alibaba
공급자입력출력
1DeepInfra최저가
$0
$0
2Alibaba주요
$0.6
$3.6

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크