GPT-5.4 (low)
OpenAIGPTProprietary
描述
GPT-5.4 is OpenAI's most capable and efficient frontier model for professional work. It combines industry-leading coding capabilities with native computer-use, up to 1M tokens of context, full-resolution vision processing, tool search for large tool ecosystems, and improved reasoning across spreadsheets, presentations, and documents. It is the most token-efficient reasoning model in the GPT-5 series.
發布日期
2026-03-05
參數規模
—
上下文長度
1.1M
支援模態
image, pdf, text
能力雷達圖
38
general
50
coding
87
reasoning
63
science
70
agents
85
multimodal
排行榜排名
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))Agents
BrowseCompOpenAI (2025)
82.7%自報
Terminal-Bench 2.0Stanford × Laude Institute (2026)
75.1%自報
OSWorld-Verified
75.0%自報
MCP Atlas
67.2%自報
SWE-Bench ProPrinceton NLP (2024)
57.7%自報
Finance Agent
56.0%自報
Toolathlon
54.6%自報
FrontierSWE
54.0%
DeepSWE 1.1
52.0%
Legal Agent Benchmark
0.4%
Biology
GPQANYU + Cohere + Anthropic (2023)
92.8%自報
Communication
Tau2 Telecom
98.9%自報
General
MMMU-Pro
81.2%自報
LiveBench
80.3%
Long Context
Graphwalks parents >128k
32.4%自報
Graphwalks BFS >128k
21.4%自報
Math
FrontierMath
47.6%自報
Humanity's Last Exam
39.8%自報
Multimodal
OmniDocBench 1.5
89.1%自報
Reasoning
ARC-AGI
93.7%自報
Graphwalks BFS <128k
93.0%自報
Graphwalks parents <128k
89.8%自報
ARC-AGI v2
73.3%自報
AA 評測指數
(Artificial Analysis)Intelligence Index(Artificial Analysis)40.2
Gpqa(NYU + Cohere + Anthropic (2023))0.9
Tau2(Sierra + U Toronto + Vector Institute (2025))0.7
Lcr(Artificial Analysis)0.7
Ifbench(Google Research (2023))0.7
Scicode(UIUC + Argonne National Lab (2024))0.5
Terminalbench Hard(Stanford × Laude Institute (2026))0.4
Hle(Center for AI Safety + Scale AI (2025))0.3
LLM Stats 分類評分
(LLM Stats (zeroeval))Communication100
Physics90
Structured Output90
Biology90
Chemistry90
Multimodal80
Search80
Vision80
Reasoning70
Spatial Reasoning70
Tool Calling70
Math60
Finance60
General60
Agents60
Code60
Long Context40
Healthcare30
Legal0
定價
輸入價格$2.5 / 1M tokens
輸出價格$15 / 1M tokens
混合價格(3:1)$5.625 / 1M tokens
快取讀取價格$0.25 / 1M tokens
速度
Tokens/秒0.0
首Token延遲0.00s
首回答延遲0.00s
供應商價格排行
供應商價格排行
1 個供應商
供應商輸入輸出
1OpenAI
$0
$0.00002
比較該模型在不同 API 供應商之間的定價。