GPT-4.1 mini
OpenAIGPTProprietary
説明
GPT-4.1 mini provides a balance between intelligence, speed, and cost. It's a significant leap in small model performance, even beating GPT-4o in many benchmarks while reducing latency and cost.
リリース日
2025-04-14
パラメータ
—
コンテキスト長
1.0M
モダリティ
image, pdf, text
能力レーダー
30
general
31
coding
54
reasoning
48
science
50
agents
85
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| コーディングランキング | 405 | 33.0 | AA |
| 総合ランキング | 345 | 40.0 | AA |
| マルチモーダルランキング | 77 | 52.0 | LS |
| 科学 | 410 | 36.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Chat
IFEvalGoogle Research (2023)
84.1%自己申告
Multi-IF
67.0%自己申告
TAU-bench Retail
55.8%自己申告
Multi-Challenge
35.8%自己申告
General
MMLU
87.5%自己申告
Internal API instruction following (hard)
45.1%自己申告
Aider-Polyglot
34.7%自己申告
Aider-Polyglot Edit
31.6%自己申告
Language
MMMLU
78.5%自己申告
COLLIE
54.6%自己申告
Long Context
ComplexFuncBench
49.3%自己申告
OpenAI-MRCR: 2 needle 128k
47.2%自己申告
OpenAI-MRCR: 2 needle 1M
33.3%自己申告
Math
MathVista
73.1%自己申告
AIME 2024
49.6%自己申告
AIME 2025
40.2%自己申告
HMMT 2025
35.0%自己申告
Multimodal
MMMU
72.7%自己申告
Reasoning
CharXiv-D
88.4%自己申告
GPQANYU + Cohere + Anthropic (2023)
65.0%自己申告
Graphwalks BFS <128k
61.7%自己申告
Graphwalks parents <128k
60.5%自己申告
CharXiv-R
56.8%自己申告
TAU-bench Airline
36.0%自己申告
SWE-Bench Verified
23.6%自己申告
Graphwalks BFS >128k
15.0%自己申告
Graphwalks parents >128k
11.0%自己申告
Humanity's Last Exam
3.7%自己申告
AA評価指数
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))92.5
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))78.1
Gpqa(NYU + Cohere + Anthropic (2023))66.4
Tau2(Sierra + U Toronto + Vector Institute (2025))52.9
Livecodebench(UC Berkeley + MIT + Cornell (2024))48.3
Aime 25(MAA (Mathematical Association of America))46.3
Math Index(Artificial Analysis)46.3
Lcr(Artificial Analysis)44.0
Aime(MAA (Mathematical Association of America))43.0
Ifbench(Google Research (2023))38.3
Coding Index(Artificial Analysis)20.2
Intelligence Index(Artificial Analysis)10.2
Terminalbench V2 110.1
Terminalbench Hard(Stanford × Laude Institute (2026))7.6
Tau Banking5.4
Hle(Center for AI Safety + Scale AI (2025))5.0
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Legal90
Finance90
Instruction Following80
Healthcare80
Language70
Multimodal70
Physics70
Structured Output70
Biology70
Chemistry70
Chat60
Vision60
Math50
Reasoning50
General50
Communication50
Tool Calling50
Writing50
Spatial Reasoning40
Long Context30
Code30
Frontend Development20
価格設定
入力価格$0.4 / 1Mトークン
出力価格$1.6 / 1Mトークン
混合価格(3:1)$0.7 / 1Mトークン
キャッシュ読み取り価格$0.1 / 1Mトークン
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
24 プロバイダー
最安: OpenAI最高: Cortecs
プロバイダー入力出力
1OpenAI最安
$0
$0
2Ofox
$0.32
$1.28
3Poe
$0.36
$1.4
4Helicone
$0.4
$1.6
5302.AI
$0.4
$1.6
6NanoGPT
$0.4
$1.6
7Abacus
$0.4
$1.6
8OpenRouter
$0.4
$1.6
9Kilo Gateway
$0.4
$1.6
10SAP AI Core
$0.4
$1.6
11Cloudflare AI Gateway
$0.4
$1.6
12Azure Cognitive Services
$0.4
$1.6
13Vercel AI Gateway
$0.4
$1.6
14DevPass (LLM Gateway)
$0.4
$1.6
15Azure
$0.4
$1.6
16NEAR AI Cloud
$0.4
$1.6
17OrcaRouter
$0.4
$1.6
18Merge Gateway
$0.4
$1.6
19Pioneer
$0.4
$1.6
20Impossibl
$0.4
$1.6
21Eden AI
$0.4
$1.6
22LLM Gateway
$0.4
$1.6
23Aixy
$0.4
$1.6
24Cortecs
$0.434
$1.704
このモデルの異なるAPIプロバイダー間の価格を比較。