GPT-5.2 (xhigh)
OpenAIGPTProprietary
説明
GPT‑5.2 introduces substantial gains in professional knowledge work, outperforming experts on GDPval with 70.9% wins or ties, and setting new highs in coding (SWE‑Bench Pro 55.6%), science (GPQA Diamond ~92–93%), math (AIME 2025: 100%), long‑context accuracy up to 256k tokens, and reliable tool‑calling (Tau2 Telecom 98.7%). It rolls out as Instant, Thinking, and Pro—faster, more structured, and less error‑prone—priced at $1.75/1M input and $14/1M output tokens, with Pro variants supporting xhigh reasoning for top‑quality, end‑to‑end execution.
リリース日
2025-12-11
パラメータ
—
コンテキスト長
400K
モダリティ
image, text
能力レーダー
56
general
81
coding
98
reasoning
66
science
70
agents
85
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| エージェント能力 | 101 | 39.0 | LS |
| コーディングランキング | 36 | 88.0 | AA |
| 総合ランキング | 36 | 82.0 | AA |
| マルチモーダルランキング | 32 | 55.0 | LS |
| 科学 | 42 | 83.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Agents
BrowseCompOpenAI (2025)
65.8%自己申告
MCP Atlas
60.6%自己申告
Toolathlon
46.3%自己申告
Biology
GPQANYU + Cohere + Anthropic (2023)
92.4%自己申告
Code
SWE-Bench Verified
80.0%自己申告
SWE-Lancer (IC-Diamond subset)
74.6%自己申告
Communication
Tau2 Telecom
98.7%自己申告
Tau2 Retail
82.0%自己申告
General
MMMLU
89.6%自己申告
MMMU-Pro
79.5%自己申告
LiveBench
74.8%
Grounding
ScreenSpot Pro
86.3%自己申告
Healthcare
VideoMMMU
85.9%自己申告
Math
AIME 2025
100.0%自己申告
HMMT 2025
99.4%自己申告
FrontierMath
40.3%自己申告
Humanity's Last Exam
34.5%自己申告
Multimodal
CharXiv-R
82.1%自己申告
Reasoning
Graphwalks BFS <128k
94.0%自己申告
BrowseComp Long Context 128k
92.0%自己申告
BrowseComp Long Context 256k
89.8%自己申告
Graphwalks parents <128k
89.0%自己申告
ARC-AGI
86.2%自己申告
ARC-AGI v2
52.9%自己申告
AA評価指数
(Artificial Analysis)Math Index(Artificial Analysis)99.0
Intelligence Index(Artificial Analysis)43.3
Aime 25(MAA (Mathematical Association of America))1.0
Gpqa(NYU + Cohere + Anthropic (2023))0.9
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.9
Tau2(Sierra + U Toronto + Vector Institute (2025))0.8
Lcr(Artificial Analysis)0.8
Ifbench(Google Research (2023))0.8
Scicode(UIUC + Argonne National Lab (2024))0.5
Terminalbench Hard(Stanford × Laude Institute (2026))0.5
Hle(Center for AI Safety + Scale AI (2025))0.4
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Physics90
Language90
Grounding90
Healthcare90
Biology90
Chemistry90
Communication90
Multimodal80
Reasoning80
Search80
Spatial Reasoning80
Frontend Development80
General80
Math70
Code70
Tool Calling70
Vision70
Agents60
価格設定
入力価格$1.75 / 1Mトークン
出力価格$14 / 1Mトークン
混合価格(3:1)$4.813 / 1Mトークン
キャッシュ読み取り価格$0.175 / 1Mトークン
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
2 プロバイダー
最安: OpenAI最高: Neon
プロバイダー入力出力
1OpenAI最安
$0
$0.00001
2Neon
$1.75
$14
このモデルの異なるAPIプロバイダー間の価格を比較。