GPT-4.1 mini
OpenAIGPTProprietary
Описание
GPT-4.1 mini provides a balance between intelligence, speed, and cost. It's a significant leap in small model performance, even beating GPT-4o in many benchmarks while reducing latency and cost.
Дата выхода
2025-04-14
Параметры
—
Длина контекста
1.0M
Модальности
image, pdf, text
Радар способностей
30
general
31
coding
54
reasoning
48
science
50
agents
85
multimodal
Рейтинги
| Домен | #Место | Оценка | Источник |
|---|---|---|---|
| Рейтинг кодинга | 405 | 33.0 | AA |
| Общий рейтинг | 345 | 40.0 | AA |
| Мультимодальный рейтинг | 77 | 52.0 | LS |
| Наука | 410 | 36.0 | AA |
Оценки бенчмарков (LLM Stats)
(LLM Stats (zeroeval))Chat
IFEvalGoogle Research (2023)
84.1%Сам.
Multi-IF
67.0%Сам.
TAU-bench Retail
55.8%Сам.
Multi-Challenge
35.8%Сам.
General
MMLU
87.5%Сам.
Internal API instruction following (hard)
45.1%Сам.
Aider-Polyglot
34.7%Сам.
Aider-Polyglot Edit
31.6%Сам.
Language
MMMLU
78.5%Сам.
COLLIE
54.6%Сам.
Long Context
ComplexFuncBench
49.3%Сам.
OpenAI-MRCR: 2 needle 128k
47.2%Сам.
OpenAI-MRCR: 2 needle 1M
33.3%Сам.
Math
MathVista
73.1%Сам.
AIME 2024
49.6%Сам.
AIME 2025
40.2%Сам.
HMMT 2025
35.0%Сам.
Multimodal
MMMU
72.7%Сам.
Reasoning
CharXiv-D
88.4%Сам.
GPQANYU + Cohere + Anthropic (2023)
65.0%Сам.
Graphwalks BFS <128k
61.7%Сам.
Graphwalks parents <128k
60.5%Сам.
CharXiv-R
56.8%Сам.
TAU-bench Airline
36.0%Сам.
SWE-Bench Verified
23.6%Сам.
Graphwalks BFS >128k
15.0%Сам.
Graphwalks parents >128k
11.0%Сам.
Humanity's Last Exam
3.7%Сам.
Индексы оценки AA
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))92.5
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))78.1
Gpqa(NYU + Cohere + Anthropic (2023))66.4
Tau2(Sierra + U Toronto + Vector Institute (2025))52.9
Livecodebench(UC Berkeley + MIT + Cornell (2024))48.3
Aime 25(MAA (Mathematical Association of America))46.3
Math Index(Artificial Analysis)46.3
Lcr(Artificial Analysis)44.0
Aime(MAA (Mathematical Association of America))43.0
Ifbench(Google Research (2023))38.3
Coding Index(Artificial Analysis)20.2
Intelligence Index(Artificial Analysis)10.2
Terminalbench V2 110.1
Terminalbench Hard(Stanford × Laude Institute (2026))7.6
Tau Banking5.4
Hle(Center for AI Safety + Scale AI (2025))5.0
Оценки категорий LLM Stats
(LLM Stats (zeroeval))Legal90
Finance90
Instruction Following80
Healthcare80
Language70
Multimodal70
Physics70
Structured Output70
Biology70
Chemistry70
Chat60
Vision60
Math50
Reasoning50
General50
Communication50
Tool Calling50
Writing50
Spatial Reasoning40
Long Context30
Code30
Frontend Development20
Цены
Цена ввода$0.4 / 1M токенов
Цена вывода$1.6 / 1M токенов
Смешанная цена (3:1)$0.7 / 1M токенов
Цена чтения кэша$0.1 / 1M токенов
Скорость
Токенов/сек0.0
Задержка первого токена0.00s
Время до первого ответа0.00s
Рейтинг цен провайдеров
Рейтинг цен провайдеров
24 провайдеров
Самый дешевый: OpenAIСамый дорогой: Cortecs
ПровайдерВводВывод
1OpenAIСамый дешевый
$0
$0
2Ofox
$0.32
$1.28
3Poe
$0.36
$1.4
4Helicone
$0.4
$1.6
5302.AI
$0.4
$1.6
6NanoGPT
$0.4
$1.6
7Abacus
$0.4
$1.6
8OpenRouter
$0.4
$1.6
9Kilo Gateway
$0.4
$1.6
10SAP AI Core
$0.4
$1.6
11Cloudflare AI Gateway
$0.4
$1.6
12Azure Cognitive Services
$0.4
$1.6
13Vercel AI Gateway
$0.4
$1.6
14DevPass (LLM Gateway)
$0.4
$1.6
15Azure
$0.4
$1.6
16NEAR AI Cloud
$0.4
$1.6
17OrcaRouter
$0.4
$1.6
18Merge Gateway
$0.4
$1.6
19Pioneer
$0.4
$1.6
20Impossibl
$0.4
$1.6
21Eden AI
$0.4
$1.6
22LLM Gateway
$0.4
$1.6
23Aixy
$0.4
$1.6
24Cortecs
$0.434
$1.704
Сравнение цен разных API-провайдеров для этой модели.