跳轉到主要內容

GPT-5.6 Terra (Max)

OpenAIGPTProprietary

描述

GPT-5.6 Terra is the balanced tier of OpenAI's GPT-5.6 family, designed for workloads that balance intelligence and cost. It delivers performance competitive with GPT-5.5 at roughly half the price, supports max reasoning effort, and has a 1.05M-token context window.

發布日期
2026-07-09
參數規模
—
上下文長度
1.1M
支援模態
image, pdf, text

能力雷達圖

42
general
73
coding
93
reasoning
69
science
60
agents
85
multimodal

排行榜排名

領域#排名分數來源
智慧體能力模型榜71
44.0
LS
程式碼能力榜22
92.0
AA
通用能力榜13
79.0
AA
數學推理13
95.0
LB
多模態榜77
52.0
LS
推理能力8
91.0
LB
科學能力24
82.0
AA

基準測試分數 (LLM Stats)

(LLM Stats (zeroeval))

Agents

Connectors100.0%自報
Search and Function-Calling94.6%自報
Capture-the-Flag Challenges (Internal)91.8%自報
Internal Research Debugging Evaluation67.8%自報
SEC-bench Pro57.7%自報
RSI Index56.3%自報
LifeSciBench56.0%自報
Toolathlon53.1%自報
PostTrainBench Lite51.5%自報
Big Finance Bench51.0%自報
Agents' Last Exam50.4%自報
KernelGen 1P49.2%自報
Management Consulting Tasks (Internal)37.2%自報
GeneBench-Pro23.3%自報
ExploitGym23.2%自報
AutomationBench15.2%自報
NanoGPT14.5%自報

Code

BenchCAD (with Python tool)78.2%自報
DeepSWE 1.170.0%
DeepSWE69.6%自報
BenchCAD62.3%自報
ExploitBench52.9%自報

General

Artificial Analysis55.0%
GDP.pdf24.7%自報

Healthcare

HealthBench Consensus95.1%自報
HealthBench Professional57.7%自報
HealthBench57.0%自報
HealthBench Hard32.7%自報

Long Context

MRCR v2 (8-needle)89.6%自報
MRCR v2 (8-needle, 512K-1M)72.5%自報

Math

FrontierMath84.9%自報
FrontierMath Tier 4 (v2)68.3%自報

Multimodal

OSWorld 2.050.2%自報

Reasoning

GPQANYU + Cohere + Anthropic (2023)92.9%自報
BrowseCompOpenAI (2025)87.5%自報
Terminal-Bench 2.187.4%自報
Graphwalks BFS >128k76.9%自報
Graphwalks BFS 1M71.2%自報
SWE-Bench ProPrinceton NLP (2024)63.4%自報
FrontierCode 1.141.3%
MedChemBench (Internal)35.0%自報
Terminal-Bench 4.021.5%
ARC-AGI-30.8%自報

Vision

MMMU-Pro (with tools)82.0%自報
MMMU-Pro80.7%自報

AA 評測指數

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
92.5
Terminalbench V2 1
88.0
Tau2(Sierra + U Toronto + Vector Institute (2025))
86.3
Lcr(Artificial Analysis)
83.0
Coding Index(Artificial Analysis)
76.7
Ifbench(Google Research (2023))
71.2
Terminalbench Hard(Stanford × Laude Institute (2026))
57.6
Scicode(UIUC + Argonne National Lab (2024))
55.0
Hle(Center for AI Safety + Scale AI (2025))
42.9
Intelligence Index(Artificial Analysis)
42.1
Tau Banking
40.2
Terminalbench V4 0
35.4

LLM Stats 分類評分

(LLM Stats (zeroeval))
Physics
90
Search
90
Biology
90
Chemistry
90
Long Context
80
Math
80
Spatial Reasoning
70
Multimodal
60
Reasoning
60
Safety
60
General
60
Healthcare
60
Agents
60
Code
60
Tool Calling
60
Vision
60
Science
40
Finance
40
Systems
40

定價

輸入價格$2 / 1M tokens
輸出價格$12 / 1M tokens
混合價格(3:1)$4.5 / 1M tokens
快取讀取價格$0.2 / 1M tokens
快取寫入價格$2.5 / 1M tokens

速度

Tokens/秒97.9
首Token延遲81.76s
首回答延遲81.76s

供應商價格排行

供應商價格排行

2 個供應商

最便宜: OpenAI最貴: Neon
供應商輸入輸出
1OpenAI最便宜
$0
$0.00001
2Neon
$2
$12

比較該模型在不同 API 供應商之間的定價。

外部連結