Skip to main content

GPT-5.6 Terra (max)

OpenAIGPTProprietary

Description

GPT-5.6 Terra is the balanced tier of OpenAI's GPT-5.6 family, designed for workloads that balance intelligence and cost. It delivers performance competitive with GPT-5.5 at roughly half the price, supports max reasoning effort, and has a 1.05M-token context window.

Release Date
2026-07-09
Parameters
Context Length
1.1M
Modalities
image, pdf, text

Capability Radar

54
general
73
coding
93
reasoning
69
science
70
agents
85
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability13
63.0
LS
Code Ranking10
95.0
AA
General Ranking16
88.0
AA
Multimodal Ranking38
57.0
LS
Science21
88.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

Connectors100.0%SR
Search and Function-Calling94.6%SR
Capture-the-Flag Challenges (Internal)91.8%SR
Internal Research Debugging Evaluation67.8%SR
SEC-bench Pro57.7%SR
RSI Index56.3%SR
LifeSciBench56.0%SR
Toolathlon53.1%SR
PostTrainBench Lite51.5%SR
Big Finance Bench51.0%SR
Agents' Last Exam50.4%SR
KernelGen 1P49.2%SR
Management Consulting Tasks (Internal)37.2%SR
GeneBench-Pro23.3%SR
ExploitGym23.2%SR
AutomationBench15.2%SR
NanoGPT14.5%SR

Code

BenchCAD (with Python tool)78.2%SR
DeepSWE 1.170.0%
DeepSWE69.6%SR
BenchCAD62.3%SR
ExploitBench52.9%SR

General

Artificial Analysis55.0%
GDP.pdf24.7%SR

Healthcare

HealthBench Consensus95.1%SR
HealthBench Professional57.7%SR
HealthBench57.0%SR
HealthBench Hard32.7%SR

Long Context

MRCR v2 (8-needle)89.6%SR
MRCR v2 (8-needle, 512K-1M)72.5%SR

Math

FrontierMath84.9%SR
FrontierMath Tier 4 (v2)68.3%SR

Multimodal

OSWorld 2.050.2%SR

Reasoning

GPQANYU + Cohere + Anthropic (2023)92.9%SR
BrowseCompOpenAI (2025)87.5%SR
Terminal-Bench 2.187.4%SR
Graphwalks BFS >128k76.9%SR
Graphwalks BFS 1M71.2%SR
SWE-Bench ProPrinceton NLP (2024)63.4%SR
FrontierCode 1.141.3%
MedChemBench (Internal)35.0%SR
ARC-AGI-30.8%SR

Vision

MMMU-Pro (with tools)82.0%SR
MMMU-Pro80.7%SR

AA Evaluation Indices

(Artificial Analysis)
Coding Index(Artificial Analysis)
76.7
Intelligence Index(Artificial Analysis)
56.6
Gpqa(NYU + Cohere + Anthropic (2023))
0.9
Terminalbench V2 1
0.9
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.9
Lcr(Artificial Analysis)
0.8
Ifbench(Google Research (2023))
0.7
Terminalbench Hard(Stanford × Laude Institute (2026))
0.6
Scicode(UIUC + Argonne National Lab (2024))
0.5
Hle(Center for AI Safety + Scale AI (2025))
0.4
Tau Banking
0.4

LLM Stats Category Scores

(LLM Stats (zeroeval))
Physics
90
Search
90
Biology
90
Chemistry
90
Long Context
80
Math
80
Spatial Reasoning
70
Tool Calling
70
Multimodal
60
Reasoning
60
Safety
60
General
60
Healthcare
60
Agents
60
Code
60
Vision
60
Science
40
Finance
40
Systems
40

Pricing

Input Price$2 / 1M tokens
Output Price$12 / 1M tokens
Blended Price (3:1)$4.5 / 1M tokens
Cache Read Price$0.2 / 1M tokens
Cache Write Price$2.5 / 1M tokens

Speed

Tokens/sec113.9
Time to First Token142.73s
Time to Answer142.73s

Provider Price Ranking

Provider Price Ranking

2 providers

Cheapest: OpenAIMost Expensive: Neon
ProviderInputOutput
1OpenAICheapest
$0
$0.00001
2Neon
$2.5
$15

Compare pricing across different API providers for this model.

External Sources