GPT-5.5

Released

GPT-5.5 is an AI model released by OpenAI on Thursday, Apr 23 2026, 37 days after GPT-5.4 nano. It has a 1.05M token context window. Benchmark results (shown below) cover Expert-SWE (Internal), GDPval (win/tie rate), Blueprint-Bench 2, BullshitBench v2, Gray Swan IPI, ProgramBench, and 23 more.

API pricing

Input
$5.00
Cached input
$0.50
Output
$30.00
  • Reasoning or thinking is supported.
USD per 1M tokens

Available from

Azure$5.00$30.001M
Azureeu$5.50$33.001M
Azureus$5.50$33.001M
Amazon Bedrockus-east-1$5.50$33.001M
USD per 1M tokens

Benchmarks

Coding

Expert-SWE (Internal)
73.1%
#1
ProgramBench
0.5%
#3 of 17
SWE-Bench Pro
58.6%
#14 of 23
SWE-Bench Multilingual
77.8%
#9 of 14
DeepSWE 1.0
64.3%
#3 of 6

Terminal & CLI

Terminal-Bench 2.1
78.2%
#20 of 30
Terminal-Bench 2.0
82.7%
#2 of 14

Agentic & tool use

MCP Atlas
75.3%
#9 of 11
Toolathlon
55.6%
#2 of 5
BrowseComp
84.4%
#12 of 28
CyberGym
81.8%
#6 of 10
OSWorld-Verified
78.7%
#9 of 28

Reasoning & science

Humanity's Last Exam
41.4%
no tools
52.2%
with tools
no tools
#9 of 22
with tools
#19 of 42
ARC-AGI-2
84.6%
#5 of 21
FrontierMath
51.7%
Tier 1–3
35.4%
Tier 4
Tier 1–3
#2 of 6
Tier 4
#3 of 6
GPQA Diamond
93.6%
#5 of 59

Knowledge work

GDPval (win/tie rate)
84.9%
#1
GDPval-AA
1769
#3 of 9
GDPval-AA v2
1494
#15 of 20

Finance

Finance Agent v2
51.8%
#5 of 9

Healthcare

MedScribe
86.87%
#4 of 5

Multimodal

Blueprint-Bench 2
36.2%
#1
CharXiv Reasoning
84.1%
#9 of 16
MMMU-Pro
81.2%
#5 of 12

Long context

MRCR v2 (8-needle)
94.8%
128k average
#2 of 7

Robustness

BullshitBench v2
47%
#33 of 77
Gray Swan IPI
3%
k = 1
17.4%
k = 10
20.8%
k = 15
k = 1
#6 of 13
k = 10
#7 of 13
k = 15
#7 of 13

Source: ProgramBench, retrieved 11 September 2026. Every other score here is the figure the lab published at launch.

About

GPT-5.5, released April 23, 2026 alongside GPT-5.5-Pro, pushed OpenAI's context window past the million-token mark (1.05M) and posted the strongest long-context recall in its class — 94.8% on MRCR v2 at 128K. Its 84.6% on ARC-AGI-2 and 51.7% on FrontierMath Tiers 1–3 led all models at release on abstract reasoning and research mathematics.

On agentic work it scored 78.2% on Terminal-Bench 2.1, 78.7% on OSWorld-Verified, and 84.9% on GDPval win-rate — the highest real-world-task result of any model at the time. GPT-5.5 headlined OpenAI's lineup for two months until the three-model GPT-5.6 family (Sol, Terra, Luna) arrived on June 26, 2026.

Compare GPT-5.5 with

GPT-5.5

Suggested comparisons

Frequently asked questions

GPT-5.5 was released by OpenAI on Thursday, Apr 23 2026.

All OpenAI releases

44 tracked

2022

1 release