Skip to main content

Mistral Small 3.2

MistralMistralOpen WeightApache 2.0 · Commercial OK

Description

Mistral-Small-3.2-24B-Instruct-2506 is a minor update of Mistral-Small-3.1-24B-Instruct-2503.

Release Date
2025-06-20
Parameters
23.6B
Context Length
128K
Modalities
image, text

Capability Radar

26
general
19
coding
40
reasoning
34
science
34
agents
90
multimodal

Rankings

Domain#RankScoreSource
Code Ranking510
19.0
AA
General Ranking450
31.0
AA
Multimodal Ranking73
53.0
LS
Science478
29.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Factuality

SimpleQA12.1%SR

General

IF84.8%SR
MMLU80.5%SR
Wild Bench65.3%SR
Arena Hard43.1%SR

Language

MMLU-Pro69.1%SR

Math

MATH69.4%SR
MathVista67.1%SR

Multimodal

MMMU62.5%SR

Reasoning

HumanEval Plus92.9%SR
ChartQAMasry et al. (2022)87.4%SR
MBPP Plus78.3%SR
GPQANYU + Cohere + Anthropic (2023)46.1%SR

Vision

DocVQADocVQA (2020)94.9%SR
AI2D92.9%SR

AA Evaluation Indices

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
88.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
68.1
Gpqa(NYU + Cohere + Anthropic (2023))
50.5
Ifbench(Google Research (2023))
33.5
Aime(MAA (Mathematical Association of America))
32.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
29.5
Scicode(UIUC + Argonne National Lab (2024))
28.6
Livecodebench(UC Berkeley + MIT + Cornell (2024))
27.5
Math Index(Artificial Analysis)
27.0
Aime 25(MAA (Mathematical Association of America))
27.0
Lcr(Artificial Analysis)
20.3
Coding Index(Artificial Analysis)
12.5
Intelligence Index(Artificial Analysis)
8.2
Terminalbench Hard(Stanford × Laude Institute (2026))
6.8
Tau Banking
6.2
Terminalbench V2 1
5.6
Hle(Center for AI Safety + Scale AI (2025))
4.3
Terminalbench V4 0
0.0

LLM Stats Category Scores

(LLM Stats (zeroeval))
Image To Text
90
Multimodal
80
Vision
80
Language
70
Legal
70
Math
70
Finance
70
General
70
Healthcare
70
Communication
70
Reasoning
60
Physics
50
Biology
50
Chemistry
50
Chat
40
Creativity
40
Writing
40
Factuality
10

Pricing

Input Price$0.075 / 1M tokens
Output Price$0.2 / 1M tokens
Blended Price (3:1)$0.106 / 1M tokens

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

3 providers

Cheapest: DeepInfraMost Expensive: DevPass (LLM Gateway)
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2MistralPRIMARY
$0.075
$0.2
3DevPass (LLM Gateway)
$0.1
$0.3

Compare pricing across different API providers for this model.

External Sources