メインコンテンツへスキップ

GPT-4.1 mini

OpenAIGPTProprietary

説明

GPT-4.1 mini provides a balance between intelligence, speed, and cost. It's a significant leap in small model performance, even beating GPT-4o in many benchmarks while reducing latency and cost.

リリース日
2025-04-14
パラメータ
—
コンテキスト長
1.0M
モダリティ
image, pdf, text

能力レーダー

30
general
31
coding
54
reasoning
48
science
50
agents
85
multimodal

ランキング

ドメイン#順位スコアソース
コーディングランキング405
33.0
AA
総合ランキング345
40.0
AA
マルチモーダルランキング77
52.0
LS
科学410
36.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Chat

IFEvalGoogle Research (2023)84.1%自己申告
Multi-IF67.0%自己申告
TAU-bench Retail55.8%自己申告
Multi-Challenge35.8%自己申告

General

MMLU87.5%自己申告
Internal API instruction following (hard)45.1%自己申告
Aider-Polyglot34.7%自己申告
Aider-Polyglot Edit31.6%自己申告

Language

MMMLU78.5%自己申告
COLLIE54.6%自己申告

Long Context

ComplexFuncBench49.3%自己申告
OpenAI-MRCR: 2 needle 128k47.2%自己申告
OpenAI-MRCR: 2 needle 1M33.3%自己申告

Math

MathVista73.1%自己申告
AIME 202449.6%自己申告
AIME 202540.2%自己申告
HMMT 202535.0%自己申告

Multimodal

MMMU72.7%自己申告

Reasoning

CharXiv-D88.4%自己申告
GPQANYU + Cohere + Anthropic (2023)65.0%自己申告
Graphwalks BFS <128k61.7%自己申告
Graphwalks parents <128k60.5%自己申告
CharXiv-R56.8%自己申告
TAU-bench Airline36.0%自己申告
SWE-Bench Verified23.6%自己申告
Graphwalks BFS >128k15.0%自己申告
Graphwalks parents >128k11.0%自己申告
Humanity's Last Exam3.7%自己申告

AA評価指数

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
92.5
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
78.1
Gpqa(NYU + Cohere + Anthropic (2023))
66.4
Tau2(Sierra + U Toronto + Vector Institute (2025))
52.9
Livecodebench(UC Berkeley + MIT + Cornell (2024))
48.3
Aime 25(MAA (Mathematical Association of America))
46.3
Math Index(Artificial Analysis)
46.3
Lcr(Artificial Analysis)
44.0
Aime(MAA (Mathematical Association of America))
43.0
Ifbench(Google Research (2023))
38.3
Coding Index(Artificial Analysis)
20.2
Intelligence Index(Artificial Analysis)
10.2
Terminalbench V2 1
10.1
Terminalbench Hard(Stanford × Laude Institute (2026))
7.6
Tau Banking
5.4
Hle(Center for AI Safety + Scale AI (2025))
5.0

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Legal
90
Finance
90
Instruction Following
80
Healthcare
80
Language
70
Multimodal
70
Physics
70
Structured Output
70
Biology
70
Chemistry
70
Chat
60
Vision
60
Math
50
Reasoning
50
General
50
Communication
50
Tool Calling
50
Writing
50
Spatial Reasoning
40
Long Context
30
Code
30
Frontend Development
20

価格設定

入力価格$0.4 / 1Mトークン
出力価格$1.6 / 1Mトークン
混合価格(3:1)$0.7 / 1Mトークン
キャッシュ読み取り価格$0.1 / 1Mトークン

速度

トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s

プロバイダー価格ランキング

プロバイダー価格ランキング

24 プロバイダー

最安: OpenAI最高: Cortecs
プロバイダー入力出力
1OpenAI最安
$0
$0
2Ofox
$0.32
$1.28
3Poe
$0.36
$1.4
4Helicone
$0.4
$1.6
5302.AI
$0.4
$1.6
6NanoGPT
$0.4
$1.6
7Abacus
$0.4
$1.6
8OpenRouter
$0.4
$1.6
9Kilo Gateway
$0.4
$1.6
10SAP AI Core
$0.4
$1.6
11Cloudflare AI Gateway
$0.4
$1.6
12Azure Cognitive Services
$0.4
$1.6
13Vercel AI Gateway
$0.4
$1.6
14DevPass (LLM Gateway)
$0.4
$1.6
15Azure
$0.4
$1.6
16NEAR AI Cloud
$0.4
$1.6
17OrcaRouter
$0.4
$1.6
18Merge Gateway
$0.4
$1.6
19Pioneer
$0.4
$1.6
20Impossibl
$0.4
$1.6
21Eden AI
$0.4
$1.6
22LLM Gateway
$0.4
$1.6
23Aixy
$0.4
$1.6
24Cortecs
$0.434
$1.704

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク