메인 콘텐츠로 건너뛰기

GPT-5.4 mini (Non-Reasoning)

OpenAIGPTProprietary

설명

GPT-5.4 mini is OpenAI's strongest mini model yet for coding, computer use, and subagents. It brings many of the strengths of GPT-5.4 to a faster, more efficient model designed for high-volume workloads, with strong coding, reasoning, multimodal understanding, and tool use.

출시일
2026-03-17
파라미터
컨텍스트 길이
400K
모달리티
image, text

능력 레이더

14
general
40
coding
61
reasoning
42
science
60
agents
85
multimodal

랭킹

도메인#순위점수소스
에이전트형 역량126
34.0
LS
코딩 랭킹322
35.0
AA
종합 랭킹450
28.0
AA
멀티모달 랭킹49
46.0
LS
과학301
45.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Agents

OSWorld-Verified72.1%자체 보고
Terminal-Bench 2.0Stanford × Laude Institute (2026)60.0%자체 보고
MCP Atlas57.7%자체 보고
SWE-Bench ProPrinceton NLP (2024)54.4%자체 보고
Finance Agent v245.4%
Toolathlon42.9%자체 보고
Legal Agent Benchmark0.0%

Biology

GPQANYU + Cohere + Anthropic (2023)88.0%자체 보고

Communication

Tau2 Telecom93.4%자체 보고

General

MMMU-Pro76.6%자체 보고
MRCR v2 (8-needle)33.6%자체 보고

Math

Humanity's Last Exam28.2%자체 보고

Multimodal

OmniDocBench 1.587.4%자체 보고

Reasoning

Graphwalks BFS <128k76.3%자체 보고
Graphwalks parents <128k71.5%자체 보고

AA 평가 지수

(Artificial Analysis)
Intelligence Index(Artificial Analysis)
16.8
Gpqa(NYU + Cohere + Anthropic (2023))
0.6
Scicode(UIUC + Argonne National Lab (2024))
0.4
Ifbench(Google Research (2023))
0.4
Lcr(Artificial Analysis)
0.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.2
Terminalbench Hard(Stanford × Laude Institute (2026))
0.2
Hle(Center for AI Safety + Scale AI (2025))
0.1

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Physics
90
Structured Output
90
Biology
90
Chemistry
90
Communication
90
Multimodal
80
Spatial Reasoning
70
Vision
70
Code
60
Tool Calling
60
Reasoning
50
Finance
50
General
50
Agents
50
Math
30
Long Context
20
Healthcare
20
Legal
0

가격

입력 가격$0.75 / 1M 토큰
출력 가격$4.5 / 1M 토큰
혼합 가격 (3:1)$1.688 / 1M 토큰
캐시 읽기 가격$0.075 / 1M 토큰

속도

토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s

공급자 가격 순위

공급자 가격 순위

1개 공급자

공급자입력출력
1OpenAI
$0
$0

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크