메인 콘텐츠로 건너뛰기

GPT-5.6 Sol (Max)

OpenAIGPTProprietary

설명

GPT-5.6 Sol is the frontier model in OpenAI's GPT-5.6 family, designed for complex professional work across coding, knowledge work, cybersecurity, and science. It sets state-of-the-art results while using fewer tokens at lower estimated cost, supports max reasoning effort and an ultra multi-agent mode, and has a 1.05M-token context window. The gpt-5.6 alias routes to GPT-5.6 Sol.

출시일
2026-07-09
파라미터
—
컨텍스트 길이
1.1M
모달리티
image, pdf, text

능력 레이더

48
general
74
coding
94
reasoning
72
science
60
agents
85
multimodal

랭킹

도메인#순위점수소스
에이전트형 역량12
65.0
LS
코딩 랭킹4
96.0
AA
종합 랭킹9
83.0
AA
수학 추론9
96.0
LB
멀티모달 랭킹39
59.0
LS
추론6
92.0
LB
과학9
87.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Agents

Connectors100.0%자체 보고
Capture-the-Flag Challenges (Internal)96.7%자체 보고
Search and Function-Calling91.0%자체 보고
SEC-bench Pro71.2%자체 보고
Internal Research Debugging Evaluation68.3%자체 보고
KernelGen 1P61.1%자체 보고
LifeSciBench59.9%자체 보고
Toolathlon58.0%자체 보고
RSI Index57.9%자체 보고
Big Finance Bench53.0%자체 보고
Agents' Last Exam52.7%자체 보고
PostTrainBench Lite50.3%자체 보고
Management Consulting Tasks (Internal)43.2%자체 보고
ExploitGym33.7%자체 보고
GeneBench-Pro28.7%자체 보고
AutomationBench18.1%자체 보고
NanoGPT9.7%자체 보고

Code

BenchCAD (with Python tool)83.4%자체 보고
ExploitBench73.5%자체 보고
DeepSWE 1.173.0%
DeepSWE72.7%자체 보고
BenchCAD70.6%자체 보고

General

Artificial Analysis59.0%
GDP.pdf30.7%자체 보고

Healthcare

HealthBench Consensus95.5%자체 보고
HealthBench Professional60.5%자체 보고
HealthBench57.0%자체 보고
HealthBench Hard33.1%자체 보고

Long Context

MRCR v2 (8-needle)91.5%자체 보고
MRCR v2 (8-needle, 512K-1M)73.8%자체 보고

Math

FrontierMath89.0%자체 보고
FrontierMath Tier 4 (v2)83.0%자체 보고

Multimodal

OSWorld 2.062.6%자체 보고

Reasoning

GPQANYU + Cohere + Anthropic (2023)94.6%자체 보고
Graphwalks BFS >128k90.7%자체 보고
BrowseCompOpenAI (2025)90.4%자체 보고
Terminal-Bench 2.188.8%자체 보고
Graphwalks BFS 1M77.1%자체 보고
SWE-Bench ProPrinceton NLP (2024)64.6%자체 보고
MedChemBench (Internal)48.3%자체 보고
FrontierCode 1.147.5%
Terminal-Bench 4.037.3%
ARC-AGI-37.8%자체 보고

Vision

MMMU-Pro (with tools)84.6%자체 보고
MMMU-Pro83.0%자체 보고

AA 평가 지수

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
94.1
Terminalbench V2 1
88.0
Tau2(Sierra + U Toronto + Vector Institute (2025))
85.1
Lcr(Artificial Analysis)
84.0
Coding Index(Artificial Analysis)
77.4
Ifbench(Google Research (2023))
72.7
Terminalbench Hard(Stanford × Laude Institute (2026))
65.9
Scicode(UIUC + Argonne National Lab (2024))
57.1
Hle(Center for AI Safety + Scale AI (2025))
49.5
Intelligence Index(Artificial Analysis)
47.0
Tau Banking
44.3
Terminalbench V4 0
39.9

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Math
90
Physics
90
Search
90
Biology
90
Chemistry
90
Long Context
80
Spatial Reasoning
80
Multimodal
70
Safety
70
Vision
70
Reasoning
60
General
60
Healthcare
60
Agents
60
Code
60
Tool Calling
60
Science
50
Finance
50
Systems
40

가격

입력 가격$4 / 1M 토큰
출력 가격$20 / 1M 토큰
혼합 가격 (3:1)$8 / 1M 토큰
캐시 읽기 가격$0.4 / 1M 토큰
캐시 쓰기 가격$5 / 1M 토큰

속도

토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s

공급자 가격 순위

공급자 가격 순위

2개 공급자

최저가: OpenAI최고가: Neon
공급자입력출력
1OpenAI최저가
$0.00001
$0.00003
2Neon
$5
$30

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크