메인 콘텐츠로 건너뛰기

GPT-5.1 (high)

OpenAIGPTProprietary

설명

The best model for coding and agentic tasks with configurable reasoning effort. GPT-5.1 is our flagship model for coding and agentic tasks with configurable reasoning and non-reasoning effort.

출시일
2025-11-13
파라미터
—
컨텍스트 길이
400K
모달리티
image, text

능력 레이더

44
general
64
coding
93
reasoning
69
science
80
agents
90
multimodal

랭킹

도메인#순위점수소스
코딩 랭킹145
75.0
AA
종합 랭킹84
69.0
AA
멀티모달 랭킹17
64.0
LS
과학137
68.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Chat

Tau2 Airline67.0%자체 보고

Communication

Tau2 Telecom95.6%자체 보고
Tau2 Retail77.9%자체 보고

Math

AIME 202594.0%자체 보고
LiveBench72.0%
FrontierMath26.7%자체 보고

Multimodal

MMMU85.4%자체 보고

Reasoning

BrowseComp Long Context 128k90.0%자체 보고
GPQANYU + Cohere + Anthropic (2023)88.1%자체 보고
SWE-Bench Verified76.3%자체 보고

AA 평가 지수

(Artificial Analysis)
Math Index(Artificial Analysis)
94.0
Aime 25(MAA (Mathematical Association of America))
94.0
Gpqa(NYU + Cohere + Anthropic (2023))
87.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
87.0
Livecodebench(UC Berkeley + MIT + Cornell (2024))
86.8
Tau2(Sierra + U Toronto + Vector Institute (2025))
81.9
Lcr(Artificial Analysis)
80.0
Ifbench(Google Research (2023))
72.9
Terminalbench V2 1
52.4
Coding Index(Artificial Analysis)
49.4
Terminalbench Hard(Stanford × Laude Institute (2026))
45.5
Hle(Center for AI Safety + Scale AI (2025))
28.5
Intelligence Index(Artificial Analysis)
24.7
Tau Banking
15.9

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Multimodal
90
Physics
90
Search
90
Healthcare
90
Biology
90
Chemistry
90
Vision
90
Reasoning
80
Frontend Development
80
General
80
Code
80
Communication
80
Tool Calling
80
Chat
70
Math
60

가격

입력 가격$1.25 / 1M 토큰
출력 가격$10 / 1M 토큰
혼합 가격 (3:1)$3.438 / 1M 토큰
캐시 읽기 가격$0.125 / 1M 토큰

속도

토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s

공급자 가격 순위

공급자 가격 순위

2개 공급자

최저가: OpenAI최고가: Neon
공급자입력출력
1OpenAI최저가
$0
$0.00001
2Neon
$1.25
$10

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크