Passer au contenu principal

QwQ 32B-Preview

AlibabaQwenOpen WeightApache 2.0 · Usage Commercial

Description

An experimental research model focused on advancing AI reasoning capabilities, particularly excelling in mathematics and programming. Features deep introspection and self-questioning abilities while having some limitations in language mixing and recursive reasoning patterns.

Date de sortie
2024-11-27
Paramètres
32.5B
Longueur du contexte
131K
Modalités
text

Radar de capacités

25
general
27
coding
61
reasoning
27
science
51
agents
0
multimodal

Classements

Domaine#RangScoreSource
Classement codage307
37.0
AA
Classement général411
32.0
AA
Science496
22.0
AA

Scores de benchmarks (LLM Stats)

(LLM Stats (zeroeval))

Biology

GPQANYU + Cohere + Anthropic (2023)65.2%Aut.

Code

LiveCodeBench50.0%Aut.

Math

MATH-50090.6%Aut.
AIME 202450.0%Aut.

Indices d'évaluation AA

(Artificial Analysis)
Intelligence Index(Artificial Analysis)
9.1
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
0.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.6
Gpqa(NYU + Cohere + Anthropic (2023))
0.6
Aime(MAA (Mathematical Association of America))
0.5
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.3
Hle(Center for AI Safety + Scale AI (2025))
0.0
Scicode(UIUC + Argonne National Lab (2024))
0.0

Scores par catégorie LLM Stats

(LLM Stats (zeroeval))
Math
70
Physics
70
Biology
70
Chemistry
70
Reasoning
60
General
60
Code
50

Tarification

Prix d'entréeGratuit
Prix de sortieGratuit
Prix mixte (3:1)Gratuit

Vitesse

Tokens/sec0.0
Délai du premier token0.00s
Temps de réponse0.00s

Classement des Prix par Fournisseur

Aucune donnée de fournisseur disponible

Sources externes