GPT-5 (high)
OpenAIGPTProprietary
説明
GPT-5 is a flagship model from OpenAI designed for coding, reasoning, and agentic tasks across domains. It is optimized for coding and agentic tasks with higher reasoning capabilities and medium speed.
リリース日
2025-08-07
パラメータ
—
コンテキスト長
400K
モダリティ
image, text
能力レーダー
43
general
56
coding
95
reasoning
68
science
80
agents
90
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| コーディングランキング | 212 | 64.0 | AA |
| 総合ランキング | 85 | 69.0 | AA |
| マルチモーダルランキング | 20 | 63.0 | LS |
| 科学 | 150 | 66.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Chat
Multi-Challenge
69.6%自己申告
Tau2 Airline
62.6%自己申告
Communication
Tau2 Telecom
96.7%自己申告
Tau2 Retail
81.1%自己申告
General
MMLU
92.5%自己申告
Aider-Polyglot
88.0%自己申告
Internal API instruction following (hard)
64.0%自己申告
LongFact Objects
0.8%自己申告
LongFact Concepts
0.7%自己申告
Healthcare
HealthBench Hard
1.6%自己申告
Language
COLLIE
99.0%自己申告
Long Context
OpenAI-MRCR: 2 needle 128k
95.2%自己申告
OpenAI-MRCR: 2 needle 256k
86.8%自己申告
Math
AIME 2025
94.6%自己申告
HMMT 2025
93.3%自己申告
MATH
84.7%自己申告
FrontierMath
26.3%自己申告
Multimodal
VideoMMMU
84.6%自己申告
MMMU
84.2%自己申告
Reasoning
SWE-Lancer (IC-Diamond subset)
100.0%自己申告
HumanEvalOpenAI (2021)
93.4%自己申告
BrowseComp Long Context 128k
90.0%自己申告
BrowseComp Long Context 256k
88.8%自己申告
GPQANYU + Cohere + Anthropic (2023)
87.3%自己申告
CharXiv-R
81.1%自己申告
Graphwalks BFS <128k
78.3%自己申告
SWE-Bench Verified
74.9%自己申告
Graphwalks parents <128k
73.3%自己申告
BrowseCompOpenAI (2025)
54.9%自己申告
Humanity's Last Exam
24.8%自己申告
FActScoreMin et al. (NYU/UW, 2023)
1.0%自己申告
Vision
VideoMME w sub.
86.7%自己申告
MMMU-Pro
78.4%自己申告
ERQA
65.7%自己申告
AA評価指数
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))99.4
Aime(MAA (Mathematical Association of America))95.7
Aime 25(MAA (Mathematical Association of America))94.3
Math Index(Artificial Analysis)94.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))87.1
Gpqa(NYU + Cohere + Anthropic (2023))85.4
Tau2(Sierra + U Toronto + Vector Institute (2025))84.8
Livecodebench(UC Berkeley + MIT + Cornell (2024))84.6
Lcr(Artificial Analysis)78.2
Ifbench(Google Research (2023))73.1
Coding Index(Artificial Analysis)37.8
Terminalbench V2 135.2
Terminalbench Hard(Stanford × Laude Institute (2026))32.6
Hle(Center for AI Safety + Scale AI (2025))28.5
Intelligence Index(Artificial Analysis)23.0
Tau Banking22.1
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Spatial Reasoning7
Vision4
Reasoning2
General1
Language100
Long Context100
Writing100
Legal90
Physics90
Finance90
Biology90
Chemistry90
Code90
Video90
Multimodal80
Communication80
Tool Calling80
Chat70
Math70
Search70
Frontend Development70
Healthcare70
Structured Output60
Agents50
価格設定
入力価格$1.25 / 1Mトークン
出力価格$10 / 1Mトークン
混合価格(3:1)$3.438 / 1Mトークン
キャッシュ読み取り価格$0.125 / 1Mトークン
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
1 プロバイダー
プロバイダー入力出力
1OpenAIプライマリ
$1.25
$10
このモデルの異なるAPIプロバイダー間の価格を比較。