GPT-5.6 Sol (max)
OpenAIGPTProprietary
説明
GPT-5.6 Sol is the frontier model in OpenAI's GPT-5.6 family, designed for complex professional work across coding, knowledge work, cybersecurity, and science. It sets state-of-the-art results while using fewer tokens at lower estimated cost, supports max reasoning effort and an ultra multi-agent mode, and has a 1.05M-token context window. The gpt-5.6 alias routes to GPT-5.6 Sol.
リリース日
2026-07-09
パラメータ
—
コンテキスト長
1.1M
モダリティ
image, pdf, text
能力レーダー
58
general
74
coding
94
reasoning
72
science
70
agents
85
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| エージェント能力 | 2 | 80.0 | LS |
| コーディングランキング | 1 | 98.0 | AA |
| 総合ランキング | 10 | 91.0 | AA |
| マルチモーダルランキング | 10 | 66.0 | LS |
| 科学 | 5 | 94.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Agents
Connectors
100.0%自己申告
Capture-the-Flag Challenges (Internal)
96.7%自己申告
Search and Function-Calling
91.0%自己申告
SEC-bench Pro
71.2%自己申告
Internal Research Debugging Evaluation
68.3%自己申告
KernelGen 1P
61.1%自己申告
LifeSciBench
59.9%自己申告
Toolathlon
58.0%自己申告
RSI Index
57.9%自己申告
Big Finance Bench
53.0%自己申告
Agents' Last Exam
52.7%自己申告
PostTrainBench Lite
50.3%自己申告
Management Consulting Tasks (Internal)
43.2%自己申告
ExploitGym
33.7%自己申告
GeneBench-Pro
28.7%自己申告
AutomationBench
18.1%自己申告
NanoGPT
9.7%自己申告
Code
BenchCAD (with Python tool)
83.4%自己申告
ExploitBench
73.5%自己申告
DeepSWE 1.1
73.0%
DeepSWE
72.7%自己申告
BenchCAD
70.6%自己申告
General
Artificial Analysis
59.0%
GDP.pdf
30.7%自己申告
Healthcare
HealthBench Consensus
95.5%自己申告
HealthBench Professional
60.5%自己申告
HealthBench
57.0%自己申告
HealthBench Hard
33.1%自己申告
Long Context
MRCR v2 (8-needle)
91.5%自己申告
MRCR v2 (8-needle, 512K-1M)
73.8%自己申告
Math
FrontierMath
89.0%自己申告
FrontierMath Tier 4 (v2)
83.0%自己申告
Multimodal
OSWorld 2.0
62.6%自己申告
Reasoning
GPQANYU + Cohere + Anthropic (2023)
94.6%自己申告
Graphwalks BFS >128k
90.7%自己申告
BrowseCompOpenAI (2025)
90.4%自己申告
Terminal-Bench 2.1
88.8%自己申告
Graphwalks BFS 1M
77.1%自己申告
SWE-Bench ProPrinceton NLP (2024)
64.6%自己申告
MedChemBench (Internal)
48.3%自己申告
FrontierCode 1.1
47.5%
ARC-AGI-3
7.8%自己申告
Vision
MMMU-Pro (with tools)
84.6%自己申告
MMMU-Pro
83.0%自己申告
AA評価指数
(Artificial Analysis)Coding Index(Artificial Analysis)77.4
Intelligence Index(Artificial Analysis)60.9
Gpqa(NYU + Cohere + Anthropic (2023))0.9
Terminalbench V2 10.9
Tau2(Sierra + U Toronto + Vector Institute (2025))0.9
Lcr(Artificial Analysis)0.8
Ifbench(Google Research (2023))0.7
Terminalbench Hard(Stanford × Laude Institute (2026))0.7
Scicode(UIUC + Argonne National Lab (2024))0.6
Hle(Center for AI Safety + Scale AI (2025))0.5
Tau Banking0.4
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Math90
Physics90
Search90
Biology90
Chemistry90
Long Context80
Spatial Reasoning80
Multimodal70
Safety70
Tool Calling70
Vision70
Reasoning60
General60
Healthcare60
Agents60
Code60
Science50
Finance50
Systems40
価格設定
入力価格$5 / 1Mトークン
出力価格$30 / 1Mトークン
混合価格(3:1)$11.25 / 1Mトークン
キャッシュ読み取り価格$0.5 / 1Mトークン
キャッシュ書き込み価格$6.25 / 1Mトークン
速度
トークン/秒74.9
初トークン遅延108.11s
初回答遅延108.11s
プロバイダー価格ランキング
プロバイダー価格ランキング
2 プロバイダー
最安: OpenAI最高: Neon
プロバイダー入力出力
1OpenAI最安
$0.00001
$0.00003
2Neon
$5
$30
このモデルの異なるAPIプロバイダー間の価格を比較。