メインコンテンツへスキップ

GPT-5.1 (high)

OpenAIGPTProprietary

説明

The best model for coding and agentic tasks with configurable reasoning effort. GPT-5.1 is our flagship model for coding and agentic tasks with configurable reasoning and non-reasoning effort.

リリース日
2025-11-13
パラメータ
—
コンテキスト長
400K
モダリティ
image, text

能力レーダー

44
general
64
coding
93
reasoning
69
science
80
agents
90
multimodal

ランキング

ドメイン#順位スコアソース
コーディングランキング145
75.0
AA
総合ランキング84
69.0
AA
マルチモーダルランキング17
64.0
LS
科学137
68.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Chat

Tau2 Airline67.0%自己申告

Communication

Tau2 Telecom95.6%自己申告
Tau2 Retail77.9%自己申告

Math

AIME 202594.0%自己申告
LiveBench72.0%
FrontierMath26.7%自己申告

Multimodal

MMMU85.4%自己申告

Reasoning

BrowseComp Long Context 128k90.0%自己申告
GPQANYU + Cohere + Anthropic (2023)88.1%自己申告
SWE-Bench Verified76.3%自己申告

AA評価指数

(Artificial Analysis)
Math Index(Artificial Analysis)
94.0
Aime 25(MAA (Mathematical Association of America))
94.0
Gpqa(NYU + Cohere + Anthropic (2023))
87.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
87.0
Livecodebench(UC Berkeley + MIT + Cornell (2024))
86.8
Tau2(Sierra + U Toronto + Vector Institute (2025))
81.9
Lcr(Artificial Analysis)
80.0
Ifbench(Google Research (2023))
72.9
Terminalbench V2 1
52.4
Coding Index(Artificial Analysis)
49.4
Terminalbench Hard(Stanford × Laude Institute (2026))
45.5
Hle(Center for AI Safety + Scale AI (2025))
28.5
Intelligence Index(Artificial Analysis)
24.7
Tau Banking
15.9

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Multimodal
90
Physics
90
Search
90
Healthcare
90
Biology
90
Chemistry
90
Vision
90
Reasoning
80
Frontend Development
80
General
80
Code
80
Communication
80
Tool Calling
80
Chat
70
Math
60

価格設定

入力価格$1.25 / 1Mトークン
出力価格$10 / 1Mトークン
混合価格(3:1)$3.438 / 1Mトークン
キャッシュ読み取り価格$0.125 / 1Mトークン

速度

トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s

プロバイダー価格ランキング

プロバイダー価格ランキング

2 プロバイダー

最安: OpenAI最高: Neon
プロバイダー入力出力
1OpenAI最安
$0
$0.00001
2Neon
$1.25
$10

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク