GPT-4.1
OpenAIGPTProprietary
説明
GPT-4.1 is OpenAI's latest and most advanced flagship model, significantly improving upon GPT-4 Turbo in performance across benchmarks, speed, and cost-effectiveness.
リリース日
2025-04-14
パラメータ
—
コンテキスト長
1.0M
モダリティ
image, pdf, text
能力レーダー
36
general
44
coding
49
reasoning
44
science
60
agents
85
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| コーディングランキング | 227 | 49.0 | AA |
| 総合ランキング | 246 | 48.0 | AA |
| マルチモーダルランキング | 53 | 43.0 | LS |
| 科学 | 291 | 45.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
66.3%自己申告
Code
SWE-Bench Verified
54.6%自己申告
Aider-Polyglot Edit
52.9%自己申告
Aider-Polyglot
51.6%自己申告
Communication
Multi-IF
70.8%自己申告
TAU-bench Retail
68.0%自己申告
TAU-bench Airline
49.4%自己申告
Multi-Challenge
38.3%自己申告
Finance
MMLU
90.2%自己申告
General
IFEvalGoogle Research (2023)
87.4%自己申告
MMMLU
87.3%自己申告
MMMU
74.8%自己申告
Internal API instruction following (hard)
49.1%自己申告
Language
COLLIE
65.8%自己申告
Long Context
ComplexFuncBench
65.5%自己申告
OpenAI-MRCR: 2 needle 128k
57.2%自己申告
OpenAI-MRCR: 2 needle 1M
46.3%自己申告
Graphwalks parents >128k
25.0%自己申告
Graphwalks BFS >128k
19.0%自己申告
Math
MathVista
72.2%自己申告
AIME 2024
48.1%自己申告
AIME 2025
46.4%自己申告
HMMT 2025
28.9%自己申告
Humanity's Last Exam
5.4%自己申告
Multimodal
CharXiv-D
87.9%自己申告
Video-MME (long, no subtitles)
72.0%自己申告
CharXiv-R
56.7%自己申告
Reasoning
Graphwalks BFS <128k
61.7%自己申告
Graphwalks parents <128k
58.0%自己申告
AA評価指数
(Artificial Analysis)Math Index(Artificial Analysis)34.7
Intelligence Index(Artificial Analysis)19.6
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))0.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.8
Gpqa(NYU + Cohere + Anthropic (2023))0.7
Lcr(Artificial Analysis)0.6
Tau2(Sierra + U Toronto + Vector Institute (2025))0.5
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.5
Aime(MAA (Mathematical Association of America))0.4
Ifbench(Google Research (2023))0.4
Scicode(UIUC + Argonne National Lab (2024))0.4
Aime 25(MAA (Mathematical Association of America))0.3
Terminalbench Hard(Stanford × Laude Institute (2026))0.1
Hle(Center for AI Safety + Scale AI (2025))0.0
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Legal90
Finance90
Instruction Following80
Language80
Healthcare80
Multimodal70
Physics70
Structured Output70
Biology70
Chemistry70
Writing70
Reasoning60
General60
Communication60
Tool Calling60
Vision60
Math50
Frontend Development50
Code50
Long Context40
Spatial Reasoning40
価格設定
入力価格$2 / 1Mトークン
出力価格$8 / 1Mトークン
混合価格(3:1)$3.5 / 1Mトークン
キャッシュ読み取り価格$0.5 / 1Mトークン
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
22 プロバイダー
最安: OpenAI最高: Cortecs
プロバイダー入力出力
1OpenAI最安
$0
$0.00001
2Poe
$1.8
$7.2
3302.AI
$2
$8
4NanoGPT
$2
$8
5Abacus
$2
$8
6OpenRouter
$2
$8
7Kilo Gateway
$2
$8
8SAP AI Core
$2
$8
9GitHub Copilot
$2
$8
10Helicone
$2
$8
11Azure Cognitive Services
$2
$8
12Vercel AI Gateway
$2
$8
13LLM Gateway
$2
$8
14Azure
$2
$8
15FastRouter
$2
$8
16NEAR AI Cloud
$2
$8
17OrcaRouter
$2
$8
18Merge Gateway
$2
$8
19Pioneer
$2
$8
20Ofox
$2
$8
21Impossibl
$2
$8
22Cortecs
$2.192
$8.769
このモデルの異なるAPIプロバイダー間の価格を比較。