GPT-4.1
OpenAIGPTProprietary
説明
GPT-4.1 is OpenAI's latest and most advanced flagship model, significantly improving upon GPT-4 Turbo in performance across benchmarks, speed, and cost-effectiveness.
リリース日
2025-04-14
パラメータ
—
コンテキスト長
1.0M
モダリティ
image, pdf, text
能力レーダー
32
general
46
coding
49
reasoning
47
science
60
agents
85
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| コーディングランキング | 306 | 49.0 | AA |
| 総合ランキング | 301 | 43.0 | AA |
| マルチモーダルランキング | 66 | 54.0 | LS |
| 科学 | 409 | 36.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Chat
IFEvalGoogle Research (2023)
87.4%自己申告
Multi-IF
70.8%自己申告
TAU-bench Retail
68.0%自己申告
Multi-Challenge
38.3%自己申告
General
MMLU
90.2%自己申告
Aider-Polyglot Edit
52.9%自己申告
Aider-Polyglot
51.6%自己申告
Internal API instruction following (hard)
49.1%自己申告
Language
MMMLU
87.3%自己申告
COLLIE
65.8%自己申告
Long Context
ComplexFuncBench
65.5%自己申告
OpenAI-MRCR: 2 needle 128k
57.2%自己申告
OpenAI-MRCR: 2 needle 1M
46.3%自己申告
Math
MathVista
72.2%自己申告
AIME 2024
48.1%自己申告
AIME 2025
46.4%自己申告
HMMT 2025
28.9%自己申告
Multimodal
MMMU
74.8%自己申告
Reasoning
CharXiv-D
87.9%自己申告
GPQANYU + Cohere + Anthropic (2023)
66.3%自己申告
Graphwalks BFS <128k
61.7%自己申告
Graphwalks parents <128k
58.0%自己申告
CharXiv-R
56.7%自己申告
SWE-Bench Verified
54.6%自己申告
TAU-bench Airline
49.4%自己申告
Graphwalks parents >128k
25.0%自己申告
Graphwalks BFS >128k
19.0%自己申告
Humanity's Last Exam
5.4%自己申告
Vision
Video-MME (long, no subtitles)
72.0%自己申告
AA評価指数
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))91.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))80.6
Lcr(Artificial Analysis)68.3
Gpqa(NYU + Cohere + Anthropic (2023))66.6
Tau2(Sierra + U Toronto + Vector Institute (2025))47.1
Livecodebench(UC Berkeley + MIT + Cornell (2024))45.7
Aime(MAA (Mathematical Association of America))43.7
Ifbench(Google Research (2023))43.0
Math Index(Artificial Analysis)34.7
Aime 25(MAA (Mathematical Association of America))34.7
Terminalbench Hard(Stanford × Laude Institute (2026))13.6
Intelligence Index(Artificial Analysis)12.7
Hle(Center for AI Safety + Scale AI (2025))4.2
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Legal90
Finance90
Instruction Following80
Language80
Healthcare80
Chat70
Multimodal70
Physics70
Structured Output70
Biology70
Chemistry70
Writing70
Reasoning60
General60
Communication60
Tool Calling60
Vision60
Math50
Frontend Development50
Code50
Long Context40
Spatial Reasoning40
価格設定
入力価格$2 / 1Mトークン
出力価格$8 / 1Mトークン
混合価格(3:1)$3.5 / 1Mトークン
キャッシュ読み取り価格$0.5 / 1Mトークン
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
24 プロバイダー
最安: OpenAI最高: Cortecs
プロバイダー入力出力
1OpenAI最安
$0
$0.00001
2Ofox
$1.6
$6.4
3Poe
$1.8
$7.2
4302.AI
$2
$8
5NanoGPT
$2
$8
6Abacus
$2
$8
7OpenRouter
$2
$8
8Kilo Gateway
$2
$8
9SAP AI Core
$2
$8
10Cloudflare AI Gateway
$2
$8
11Helicone
$2
$8
12Azure Cognitive Services
$2
$8
13Vercel AI Gateway
$2
$8
14DevPass (LLM Gateway)
$2
$8
15Azure
$2
$8
16FastRouter
$2
$8
17NEAR AI Cloud
$2
$8
18OrcaRouter
$2
$8
19Merge Gateway
$2
$8
20Pioneer
$2
$8
21Impossibl
$2
$8
22Eden AI
$2
$8
23LLM Gateway
$2
$8
24Cortecs
$2.192
$8.769
このモデルの異なるAPIプロバイダー間の価格を比較。