Перейти к основному содержанию

Qwen3 235B A22B 2507 Instruct

AlibabaQwenОткрытые весаApache 2.0 · Коммерческое использование

Описание

Qwen3-235B-A22B-Instruct-2507 is the updated instruct version of Qwen3-235B-A22B featuring significant improvements in general capabilities including instruction following, logical reasoning, text comprehension, mathematics, science, coding and tool usage. It provides substantial gains in long-tail knowledge coverage across multiple languages and markedly better alignment with user preferences in subjective and open-ended tasks.

Дата выхода
2025-07-21
Параметры
235.0B
Длина контекста
262K
Модальности
text

Радар способностей

37
general
49
coding
76
reasoning
49
science
60
agents
0
multimodal

Рейтинги

Домен#МестоОценкаИсточник
Агентные возможности53
52.0
LS
Рейтинг кодинга292
40.0
AA
Общий рейтинг280
46.0
AA
Наука232
52.0
AA

Оценки бенчмарков (LLM Stats)

(LLM Stats (zeroeval))

Agents

BFCL-v370.9%Сам.

Biology

GPQANYU + Cohere + Anthropic (2023)77.5%Сам.

Chemistry

SuperGPQA62.6%Сам.

Code

Aider-Polyglot57.3%Сам.

Communication

WritingBench85.2%Сам.
Multi-IF77.5%Сам.
Tau2 Retail71.3%Сам.
Tau2 Airline44.0%Сам.

Creativity

Creative Writing v387.5%Сам.
Arena-Hard v279.2%Сам.

Factuality

SimpleQA54.3%Сам.

Finance

MMLU-Pro83.0%Сам.
MMLU-ProX79.4%Сам.

General

MMLU-Redux93.1%Сам.
IFEvalGoogle Research (2023)88.7%Сам.
MultiPL-E87.9%Сам.
CSimpleQA84.3%Сам.
Include79.5%Сам.
LiveBench 2024112575.4%Сам.
LiveCodeBench v651.8%Сам.

Math

AIME 202570.3%Сам.
HMMT2555.4%Сам.
PolyMATH50.2%Сам.

Reasoning

ZebraLogic95.0%Сам.
ARC-AGI41.8%Сам.

Индексы оценки AA

(Artificial Analysis)
Math Index(Artificial Analysis)
71.7
Intelligence Index(Artificial Analysis)
18.4
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
1.0
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.8
Gpqa(NYU + Cohere + Anthropic (2023))
0.8
Aime(MAA (Mathematical Association of America))
0.7
Aime 25(MAA (Mathematical Association of America))
0.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.5
Ifbench(Google Research (2023))
0.5
Scicode(UIUC + Argonne National Lab (2024))
0.4
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.3
Lcr(Artificial Analysis)
0.3
Terminalbench Hard(Stanford × Laude Institute (2026))
0.2
Hle(Center for AI Safety + Scale AI (2025))
0.1

Оценки категорий LLM Stats

(LLM Stats (zeroeval))
Legal
80
Structured Output
80
Instruction Following
80
Language
80
Finance
80
Healthcare
80
Biology
80
Creativity
80
Writing
80
Math
70
Physics
70
Reasoning
70
General
70
Agents
70
Chemistry
70
Communication
70
Code
60
Economics
60
Tool Calling
60
Multimodal
50
Spatial Reasoning
50
Factuality
50
Vision
50

Цены

Цена ввода$0.7 / 1M токенов
Цена вывода$2.8 / 1M токенов
Смешанная цена (3:1)$1.225 / 1M токенов

Скорость

Токенов/сек0.0
Задержка первого токена0.00s
Время до первого ответа0.00s

Рейтинг цен провайдеров

Рейтинг цен провайдеров

7 провайдеров

Самый дешевый: CortecsСамый дорогой: GreenPT
ПровайдерВводВывод
1CortecsСамый дешевый
$0.069
$0.455
2LLM Gateway
$0.09
$0.58
3Venice AI
$0.15
$0.75
4302.AI
$0.29
$1.143
5AlibabaОсновной
$0.7
$2.8
6Scaleway
$0.75
$2.25
7GreenPT
$1.026
$3.078

Сравнение цен разных API-провайдеров для этой модели.

Внешние ссылки