GPT-5.5 (Xhigh)
OpenAIGPTProprietary
Descripción
GPT-5.5 is OpenAI's smartest model yet, designed for real work across agentic coding, computer use, knowledge work, and early scientific research. It matches GPT-5.4 per-token latency in real-world serving while reaching a much higher level of intelligence and using significantly fewer tokens to complete the same tasks. GPT-5.5 supports a 1M-token context window in the API and a 400K-token context window in Codex, with state-of-the-art results on Terminal-Bench 2.0, OSWorld-Verified, GDPval, FrontierMath, and CyberGym.
Fecha de lanzamiento
2026-04-23
Parámetros
—
Longitud del contexto
1.1M
Modalidades
image, pdf, text
Radar de capacidades
40
general
72
coding
93
reasoning
71
science
80
agents
85
multimodal
Rankings
| Dominio | #Posición | Puntuación | Fuente |
|---|---|---|---|
| Capacidad agéntica | 38 | 50.0 | LS |
| Ranking de codificación | 34 | 93.0 | AA |
| Ranking general | 34 | 79.0 | AA |
| Ranking multimodal | 21 | 62.0 | LS |
| Ciencia | 35 | 84.0 | AA |
Puntuaciones de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Agents
BixBench
80.5%Aut.
Finance Agent
60.0%Aut.
Toolathlon
55.6%Aut.
OfficeQA Pro
54.1%Aut.
Finance Agent v2
51.8%
GeneBench
25.0%Aut.
Code
CyberGym
81.8%Aut.
FrontierSWE
73.0%
DeepSWE 1.1
67.0%
Communication
Tau2 Telecom
98.0%Aut.
General
GDPval-MM
84.9%Aut.
Legal
Legal Agent Benchmark
2.1%
Long Context
MRCR v2 (8-needle)
74.0%Aut.
Math
LiveBench
80.7%
FrontierMath
35.4%Aut.
Multimodal
OSWorld-Verified
78.7%Aut.
Reasoning
ARC-AGI
95.0%Aut.
GPQANYU + Cohere + Anthropic (2023)
93.6%Aut.
ARC-AGI v2
85.0%Aut.
BrowseCompOpenAI (2025)
84.4%Aut.
Terminal-Bench 2.0Stanford × Laude Institute (2026)
82.7%Aut.
MCP Atlas
75.3%Aut.
SWE-Bench ProPrinceton NLP (2024)
58.6%Aut.
Graphwalks parents >128k
58.5%Aut.
Humanity's Last Exam
52.2%Aut.
Graphwalks BFS >128k
45.4%Aut.
FrontierCode 1.1
43.0%
Vision
MMMU-Pro
83.2%Aut.
Índices de evaluación AA
(Artificial Analysis)Tau2(Sierra + U Toronto + Vector Institute (2025))93.9
Gpqa(NYU + Cohere + Anthropic (2023))93.5
Lcr(Artificial Analysis)84.3
Terminalbench V2 184.3
Ifbench(Google Research (2023))75.9
Coding Index(Artificial Analysis)74.9
Terminalbench Hard(Stanford × Laude Institute (2026))60.6
Scicode(UIUC + Argonne National Lab (2024))55.8
Hle(Center for AI Safety + Scale AI (2025))45.8
Tau Banking39.0
Intelligence Index(Artificial Analysis)38.4
Terminalbench V4 014.6
Puntuaciones por categoría LLM Stats
(LLM Stats (zeroeval))Communication100
Physics90
Biology90
Chemistry90
Multimodal80
Safety80
Search80
Tool Calling80
Vision80
Reasoning70
Spatial Reasoning70
Finance70
General70
Code70
Long Context60
Math60
Agents60
Legal0
Precios
Precio de entrada$5 / 1M tokens
Precio de salida$30 / 1M tokens
Precio mixto (3:1)$11.25 / 1M tokens
Precio de lectura caché$0.5 / 1M tokens
Velocidad
Tokens/seg0.0
Retraso del primer token0.00s
Tiempo hasta la respuesta0.00s
Ranking de Precios por Proveedor
Ranking de Precios por Proveedor
3 proveedores
Más barato: OpenAIMás caro: Neon
ProveedorEntradaSalida
1OpenAIMás barato
$0.00001
$0.00003
2FrogBot
$2.5
$15
3Neon
$5
$30
Comparar precios entre diferentes proveedores de API para este modelo.