GPT-4o (Aug '24)
OpenAIGPTProprietary
설명
GPT-4o ('o' for 'omni') is a multimodal AI model that accepts text, audio, image, and video inputs, and generates text, audio, and image outputs. It matches GPT-4 Turbo performance on text and code, with improvements in non-English languages, vision, and audio understanding.
출시일
2024-08-06
파라미터
—
컨텍스트 길이
128K
모달리티
image, text
능력 레이더
7
general
32
coding
40
reasoning
37
science
50
agents
90
multimodal
랭킹
벤치마크 점수 (LLM Stats)
(LLM Stats (zeroeval))Chat
IFEvalGoogle Research (2023)
81.0%자체 보고
Multi-IF
60.9%자체 보고
TAU-bench Retail
60.3%자체 보고
Tau2 Airline
45.5%자체 보고
Multi-Challenge
40.3%자체 보고
Communication
Tau2 Retail
63.4%자체 보고
Tau2 Telecom
23.5%자체 보고
Factuality
SimpleQA
38.2%자체 보고
General
MMLU
85.7%자체 보고
Aider-Polyglot
30.7%자체 보고
Internal API instruction following (hard)
29.2%자체 보고
Aider-Polyglot Edit
18.2%자체 보고
Language
MMMLU
81.4%자체 보고
MMLU-Pro
74.7%자체 보고
COLLIE
61.0%자체 보고
Long Context
ComplexFuncBench
66.5%자체 보고
OpenAI-MRCR: 2 needle 128k
31.9%자체 보고
Math
MathVista
61.4%자체 보고
AIME 2024
13.1%자체 보고
Multimodal
MMMU
72.2%자체 보고
VideoMMMU
61.2%자체 보고
Reasoning
ChartQAMasry et al. (2022)
85.7%자체 보고
CharXiv-D
85.3%자체 보고
GPQANYU + Cohere + Anthropic (2023)
70.1%자체 보고
CharXiv-R
58.8%자체 보고
TAU-bench Airline
42.8%자체 보고
Graphwalks BFS <128k
41.7%자체 보고
Graphwalks parents <128k
35.4%자체 보고
SWE-Bench Verified
33.2%자체 보고
SWE-Lancer
32.6%자체 보고
SWE-Lancer (IC-Diamond subset)
12.4%자체 보고
Humanity's Last Exam
5.3%자체 보고
Vision
AI2D
94.2%자체 보고
DocVQADocVQA (2020)
92.8%자체 보고
EgoSchema
72.2%자체 보고
ActivityNet
61.9%자체 보고
MMMU-Pro
59.9%자체 보고
ERQA
35.2%자체 보고
AA 평가 지수
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))79.5
Gpqa(NYU + Cohere + Anthropic (2023))52.1
Lcr(Artificial Analysis)41.0
Ifbench(Google Research (2023))36.0
Livecodebench(UC Berkeley + MIT + Cornell (2024))31.7
Tau2(Sierra + U Toronto + Vector Institute (2025))28.9
Aime(MAA (Mathematical Association of America))11.7
Terminalbench Hard(Stanford × Laude Institute (2026))8.3
Intelligence Index(Artificial Analysis)7.7
Hle(Center for AI Safety + Scale AI (2025))2.3
LLM Stats 카테고리 점수
(LLM Stats (zeroeval))Image To Text90
Legal80
Finance80
Instruction Following70
Language70
Multimodal70
Physics70
Healthcare70
Biology70
Chemistry70
Vision70
Chat60
Long Context60
Structured Output60
Writing60
Math50
Reasoning50
General50
Communication50
Tool Calling50
Spatial Reasoning40
Factuality40
Frontend Development30
Code30
가격
입력 가격$2.5 / 1M 토큰
출력 가격$10 / 1M 토큰
혼합 가격 (3:1)$4.375 / 1M 토큰
속도
토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s
공급자 가격 순위
공급자 가격 순위
2개 공급자
최저가: OpenAI최고가: Azure
공급자입력출력
1OpenAI최저가
$0
$0.00001
2Azure
$0
$0.00001
이 모델의 다양한 API 공급자 간 가격 비교.