GPT-5.6 Terra (max)
OpenAIGPTProprietary
Description
GPT-5.6 Terra is the balanced tier of OpenAI's GPT-5.6 family, designed for workloads that balance intelligence and cost. It delivers performance competitive with GPT-5.5 at roughly half the price, supports max reasoning effort, and has a 1.05M-token context window.
Release Date
2026-07-09
Parameters
—
Context Length
1.1M
Modalities
image, pdf, text
Capability Radar
54
general
73
coding
93
reasoning
69
science
70
agents
85
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 13 | 63.0 | LS |
| Code Ranking | 10 | 95.0 | AA |
| General Ranking | 16 | 88.0 | AA |
| Multimodal Ranking | 38 | 57.0 | LS |
| Science | 21 | 88.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
Connectors
100.0%SR
Search and Function-Calling
94.6%SR
Capture-the-Flag Challenges (Internal)
91.8%SR
Internal Research Debugging Evaluation
67.8%SR
SEC-bench Pro
57.7%SR
RSI Index
56.3%SR
LifeSciBench
56.0%SR
Toolathlon
53.1%SR
PostTrainBench Lite
51.5%SR
Big Finance Bench
51.0%SR
Agents' Last Exam
50.4%SR
KernelGen 1P
49.2%SR
Management Consulting Tasks (Internal)
37.2%SR
GeneBench-Pro
23.3%SR
ExploitGym
23.2%SR
AutomationBench
15.2%SR
NanoGPT
14.5%SR
Code
BenchCAD (with Python tool)
78.2%SR
DeepSWE 1.1
70.0%
DeepSWE
69.6%SR
BenchCAD
62.3%SR
ExploitBench
52.9%SR
General
Artificial Analysis
55.0%
GDP.pdf
24.7%SR
Healthcare
HealthBench Consensus
95.1%SR
HealthBench Professional
57.7%SR
HealthBench
57.0%SR
HealthBench Hard
32.7%SR
Long Context
MRCR v2 (8-needle)
89.6%SR
MRCR v2 (8-needle, 512K-1M)
72.5%SR
Math
FrontierMath
84.9%SR
FrontierMath Tier 4 (v2)
68.3%SR
Multimodal
OSWorld 2.0
50.2%SR
Reasoning
GPQANYU + Cohere + Anthropic (2023)
92.9%SR
BrowseCompOpenAI (2025)
87.5%SR
Terminal-Bench 2.1
87.4%SR
Graphwalks BFS >128k
76.9%SR
Graphwalks BFS 1M
71.2%SR
SWE-Bench ProPrinceton NLP (2024)
63.4%SR
FrontierCode 1.1
41.3%
MedChemBench (Internal)
35.0%SR
ARC-AGI-3
0.8%SR
Vision
MMMU-Pro (with tools)
82.0%SR
MMMU-Pro
80.7%SR
AA Evaluation Indices
(Artificial Analysis)Coding Index(Artificial Analysis)76.7
Intelligence Index(Artificial Analysis)56.6
Gpqa(NYU + Cohere + Anthropic (2023))0.9
Terminalbench V2 10.9
Tau2(Sierra + U Toronto + Vector Institute (2025))0.9
Lcr(Artificial Analysis)0.8
Ifbench(Google Research (2023))0.7
Terminalbench Hard(Stanford × Laude Institute (2026))0.6
Scicode(UIUC + Argonne National Lab (2024))0.5
Hle(Center for AI Safety + Scale AI (2025))0.4
Tau Banking0.4
LLM Stats Category Scores
(LLM Stats (zeroeval))Physics90
Search90
Biology90
Chemistry90
Long Context80
Math80
Spatial Reasoning70
Tool Calling70
Multimodal60
Reasoning60
Safety60
General60
Healthcare60
Agents60
Code60
Vision60
Science40
Finance40
Systems40
Pricing
Input Price$2 / 1M tokens
Output Price$12 / 1M tokens
Blended Price (3:1)$4.5 / 1M tokens
Cache Read Price$0.2 / 1M tokens
Cache Write Price$2.5 / 1M tokens
Speed
Tokens/sec113.9
Time to First Token142.73s
Time to Answer142.73s
Provider Price Ranking
Provider Price Ranking
2 providers
Cheapest: OpenAIMost Expensive: Neon
ProviderInputOutput
1OpenAICheapest
$0
$0.00001
2Neon
$2.5
$15
Compare pricing across different API providers for this model.