GPT-4.1 mini
OpenAIGPTProprietary
Description
GPT-4.1 mini provides a balance between intelligence, speed, and cost. It's a significant leap in small model performance, even beating GPT-4o in many benchmarks while reducing latency and cost.
Date de sortie
2025-04-14
Paramètres
—
Longueur du contexte
1.0M
Modalités
image, pdf, text
Radar de capacités
32
general
32
coding
54
reasoning
45
science
50
agents
85
multimodal
Classements
| Domaine | #Rang | Score | Source |
|---|---|---|---|
| Classement codage | 325 | 34.0 | AA |
| Classement général | 295 | 44.0 | AA |
| Classement multimodal | 56 | 44.0 | LS |
| Science | 276 | 47.0 | AA |
Scores de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
65.0%Aut.
Code
Aider-Polyglot
34.7%Aut.
Aider-Polyglot Edit
31.6%Aut.
SWE-Bench Verified
23.6%Aut.
Communication
Multi-IF
67.0%Aut.
TAU-bench Retail
55.8%Aut.
TAU-bench Airline
36.0%Aut.
Multi-Challenge
35.8%Aut.
Finance
MMLU
87.5%Aut.
General
IFEvalGoogle Research (2023)
84.1%Aut.
MMMLU
78.5%Aut.
MMMU
72.7%Aut.
Internal API instruction following (hard)
45.1%Aut.
Language
COLLIE
54.6%Aut.
Long Context
ComplexFuncBench
49.3%Aut.
OpenAI-MRCR: 2 needle 128k
47.2%Aut.
OpenAI-MRCR: 2 needle 1M
33.3%Aut.
Graphwalks BFS >128k
15.0%Aut.
Graphwalks parents >128k
11.0%Aut.
Math
MathVista
73.1%Aut.
AIME 2024
49.6%Aut.
AIME 2025
40.2%Aut.
HMMT 2025
35.0%Aut.
Humanity's Last Exam
3.7%Aut.
Multimodal
CharXiv-D
88.4%Aut.
CharXiv-R
56.8%Aut.
Reasoning
Graphwalks BFS <128k
61.7%Aut.
Graphwalks parents <128k
60.5%Aut.
Indices d'évaluation AA
(Artificial Analysis)Math Index(Artificial Analysis)46.3
Coding Index(Artificial Analysis)20.2
Intelligence Index(Artificial Analysis)14.8
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))0.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.8
Gpqa(NYU + Cohere + Anthropic (2023))0.7
Tau2(Sierra + U Toronto + Vector Institute (2025))0.5
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.5
Aime 25(MAA (Mathematical Association of America))0.5
Lcr(Artificial Analysis)0.5
Aime(MAA (Mathematical Association of America))0.4
Scicode(UIUC + Argonne National Lab (2024))0.4
Ifbench(Google Research (2023))0.4
Terminalbench V2 10.1
Terminalbench Hard(Stanford × Laude Institute (2026))0.1
Tau Banking0.1
Hle(Center for AI Safety + Scale AI (2025))0.1
Scores par catégorie LLM Stats
(LLM Stats (zeroeval))Legal90
Finance90
Instruction Following80
Healthcare80
Multimodal70
Physics70
Structured Output70
Language70
Biology70
Chemistry70
Vision60
Math50
Reasoning50
General50
Communication50
Tool Calling50
Writing50
Spatial Reasoning40
Long Context30
Code30
Frontend Development20
Tarification
Prix d'entrée$0.4 / 1M tokens
Prix de sortie$1.6 / 1M tokens
Prix mixte (3:1)$0.7 / 1M tokens
Prix de lecture cache$0.1 / 1M tokens
Vitesse
Tokens/sec0.0
Délai du premier token0.00s
Temps de réponse0.00s
Classement des Prix par Fournisseur
Classement des Prix par Fournisseur
20 fournisseurs
Moins cher: OpenAIPlus cher: Cortecs
FournisseurEntréeSortie
1OpenAIMoins cher
$0
$0
2Poe
$0.36
$1.4
3Helicone
$0.4
$1.6
4302.AI
$0.4
$1.6
5NanoGPT
$0.4
$1.6
6Abacus
$0.4
$1.6
7OpenRouter
$0.4
$1.6
8Kilo Gateway
$0.4
$1.6
9SAP AI Core
$0.4
$1.6
10Azure Cognitive Services
$0.4
$1.6
11Vercel AI Gateway
$0.4
$1.6
12LLM Gateway
$0.4
$1.6
13Azure
$0.4
$1.6
14NEAR AI Cloud
$0.4
$1.6
15OrcaRouter
$0.4
$1.6
16Merge Gateway
$0.4
$1.6
17Pioneer
$0.4
$1.6
18Ofox
$0.4
$1.6
19Impossibl
$0.4
$1.6
20Cortecs
$0.434
$1.704
Comparer les prix entre différents fournisseurs API pour ce modèle.