GPT-4.1 mini
OpenAIGPTProprietary
Descripción
GPT-4.1 mini provides a balance between intelligence, speed, and cost. It's a significant leap in small model performance, even beating GPT-4o in many benchmarks while reducing latency and cost.
Fecha de lanzamiento
2025-04-14
Parámetros
—
Longitud del contexto
1.0M
Modalidades
image, pdf, text
Radar de capacidades
30
general
31
coding
54
reasoning
48
science
50
agents
85
multimodal
Rankings
| Dominio | #Posición | Puntuación | Fuente |
|---|---|---|---|
| Ranking de codificación | 405 | 33.0 | AA |
| Ranking general | 345 | 40.0 | AA |
| Ranking multimodal | 77 | 52.0 | LS |
| Ciencia | 410 | 36.0 | AA |
Puntuaciones de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Chat
IFEvalGoogle Research (2023)
84.1%Aut.
Multi-IF
67.0%Aut.
TAU-bench Retail
55.8%Aut.
Multi-Challenge
35.8%Aut.
General
MMLU
87.5%Aut.
Internal API instruction following (hard)
45.1%Aut.
Aider-Polyglot
34.7%Aut.
Aider-Polyglot Edit
31.6%Aut.
Language
MMMLU
78.5%Aut.
COLLIE
54.6%Aut.
Long Context
ComplexFuncBench
49.3%Aut.
OpenAI-MRCR: 2 needle 128k
47.2%Aut.
OpenAI-MRCR: 2 needle 1M
33.3%Aut.
Math
MathVista
73.1%Aut.
AIME 2024
49.6%Aut.
AIME 2025
40.2%Aut.
HMMT 2025
35.0%Aut.
Multimodal
MMMU
72.7%Aut.
Reasoning
CharXiv-D
88.4%Aut.
GPQANYU + Cohere + Anthropic (2023)
65.0%Aut.
Graphwalks BFS <128k
61.7%Aut.
Graphwalks parents <128k
60.5%Aut.
CharXiv-R
56.8%Aut.
TAU-bench Airline
36.0%Aut.
SWE-Bench Verified
23.6%Aut.
Graphwalks BFS >128k
15.0%Aut.
Graphwalks parents >128k
11.0%Aut.
Humanity's Last Exam
3.7%Aut.
Índices de evaluación AA
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))92.5
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))78.1
Gpqa(NYU + Cohere + Anthropic (2023))66.4
Tau2(Sierra + U Toronto + Vector Institute (2025))52.9
Livecodebench(UC Berkeley + MIT + Cornell (2024))48.3
Aime 25(MAA (Mathematical Association of America))46.3
Math Index(Artificial Analysis)46.3
Lcr(Artificial Analysis)44.0
Aime(MAA (Mathematical Association of America))43.0
Ifbench(Google Research (2023))38.3
Coding Index(Artificial Analysis)20.2
Intelligence Index(Artificial Analysis)10.2
Terminalbench V2 110.1
Terminalbench Hard(Stanford × Laude Institute (2026))7.6
Tau Banking5.4
Hle(Center for AI Safety + Scale AI (2025))5.0
Puntuaciones por categoría LLM Stats
(LLM Stats (zeroeval))Legal90
Finance90
Instruction Following80
Healthcare80
Language70
Multimodal70
Physics70
Structured Output70
Biology70
Chemistry70
Chat60
Vision60
Math50
Reasoning50
General50
Communication50
Tool Calling50
Writing50
Spatial Reasoning40
Long Context30
Code30
Frontend Development20
Precios
Precio de entrada$0.4 / 1M tokens
Precio de salida$1.6 / 1M tokens
Precio mixto (3:1)$0.7 / 1M tokens
Precio de lectura caché$0.1 / 1M tokens
Velocidad
Tokens/seg0.0
Retraso del primer token0.00s
Tiempo hasta la respuesta0.00s
Ranking de Precios por Proveedor
Ranking de Precios por Proveedor
24 proveedores
Más barato: OpenAIMás caro: Cortecs
ProveedorEntradaSalida
1OpenAIMás barato
$0
$0
2Ofox
$0.32
$1.28
3Poe
$0.36
$1.4
4Helicone
$0.4
$1.6
5302.AI
$0.4
$1.6
6NanoGPT
$0.4
$1.6
7Abacus
$0.4
$1.6
8OpenRouter
$0.4
$1.6
9Kilo Gateway
$0.4
$1.6
10SAP AI Core
$0.4
$1.6
11Cloudflare AI Gateway
$0.4
$1.6
12Azure Cognitive Services
$0.4
$1.6
13Vercel AI Gateway
$0.4
$1.6
14DevPass (LLM Gateway)
$0.4
$1.6
15Azure
$0.4
$1.6
16NEAR AI Cloud
$0.4
$1.6
17OrcaRouter
$0.4
$1.6
18Merge Gateway
$0.4
$1.6
19Pioneer
$0.4
$1.6
20Impossibl
$0.4
$1.6
21Eden AI
$0.4
$1.6
22LLM Gateway
$0.4
$1.6
23Aixy
$0.4
$1.6
24Cortecs
$0.434
$1.704
Comparar precios entre diferentes proveedores de API para este modelo.