メインコンテンツへスキップ

GPT-5.6 Sol (max)

OpenAIGPTProprietary

説明

GPT-5.6 Sol is the frontier model in OpenAI's GPT-5.6 family, designed for complex professional work across coding, knowledge work, cybersecurity, and science. It sets state-of-the-art results while using fewer tokens at lower estimated cost, supports max reasoning effort and an ultra multi-agent mode, and has a 1.05M-token context window. The gpt-5.6 alias routes to GPT-5.6 Sol.

リリース日
2026-07-09
パラメータ
コンテキスト長
1.1M
モダリティ
image, pdf, text

能力レーダー

58
general
74
coding
94
reasoning
72
science
70
agents
85
multimodal

ランキング

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Agents

Connectors100.0%自己申告
Capture-the-Flag Challenges (Internal)96.7%自己申告
Search and Function-Calling91.0%自己申告
SEC-bench Pro71.2%自己申告
Internal Research Debugging Evaluation68.3%自己申告
KernelGen 1P61.1%自己申告
LifeSciBench59.9%自己申告
Toolathlon58.0%自己申告
RSI Index57.9%自己申告
Big Finance Bench53.0%自己申告
Agents' Last Exam52.7%自己申告
PostTrainBench Lite50.3%自己申告
Management Consulting Tasks (Internal)43.2%自己申告
ExploitGym33.7%自己申告
GeneBench-Pro28.7%自己申告
AutomationBench18.1%自己申告
NanoGPT9.7%自己申告

Code

BenchCAD (with Python tool)83.4%自己申告
ExploitBench73.5%自己申告
DeepSWE 1.173.0%
DeepSWE72.7%自己申告
BenchCAD70.6%自己申告

General

Artificial Analysis59.0%
GDP.pdf30.7%自己申告

Healthcare

HealthBench Consensus95.5%自己申告
HealthBench Professional60.5%自己申告
HealthBench57.0%自己申告
HealthBench Hard33.1%自己申告

Long Context

MRCR v2 (8-needle)91.5%自己申告
MRCR v2 (8-needle, 512K-1M)73.8%自己申告

Math

FrontierMath89.0%自己申告
FrontierMath Tier 4 (v2)83.0%自己申告

Multimodal

OSWorld 2.062.6%自己申告

Reasoning

GPQANYU + Cohere + Anthropic (2023)94.6%自己申告
Graphwalks BFS >128k90.7%自己申告
BrowseCompOpenAI (2025)90.4%自己申告
Terminal-Bench 2.188.8%自己申告
Graphwalks BFS 1M77.1%自己申告
SWE-Bench ProPrinceton NLP (2024)64.6%自己申告
MedChemBench (Internal)48.3%自己申告
FrontierCode 1.147.5%
ARC-AGI-37.8%自己申告

Vision

MMMU-Pro (with tools)84.6%自己申告
MMMU-Pro83.0%自己申告

AA評価指数

(Artificial Analysis)
Coding Index(Artificial Analysis)
77.4
Intelligence Index(Artificial Analysis)
60.9
Gpqa(NYU + Cohere + Anthropic (2023))
0.9
Terminalbench V2 1
0.9
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.9
Lcr(Artificial Analysis)
0.8
Ifbench(Google Research (2023))
0.7
Terminalbench Hard(Stanford × Laude Institute (2026))
0.7
Scicode(UIUC + Argonne National Lab (2024))
0.6
Hle(Center for AI Safety + Scale AI (2025))
0.5
Tau Banking
0.4

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Math
90
Physics
90
Search
90
Biology
90
Chemistry
90
Long Context
80
Spatial Reasoning
80
Multimodal
70
Safety
70
Tool Calling
70
Vision
70
Reasoning
60
General
60
Healthcare
60
Agents
60
Code
60
Science
50
Finance
50
Systems
40

価格設定

入力価格$5 / 1Mトークン
出力価格$30 / 1Mトークン
混合価格(3:1)$11.25 / 1Mトークン
キャッシュ読み取り価格$0.5 / 1Mトークン
キャッシュ書き込み価格$6.25 / 1Mトークン

速度

トークン/秒74.9
初トークン遅延108.11s
初回答遅延108.11s

プロバイダー価格ランキング

プロバイダー価格ランキング

2 プロバイダー

最安: OpenAI最高: Neon
プロバイダー入力出力
1OpenAI最安
$0.00001
$0.00003
2Neon
$5
$30

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク