GPT-5.5-Pro
Released
GPT-5.5-Pro is an AI model released by OpenAI on Thursday, Apr 23 2026, the same day as GPT-5.5. Benchmark results (shown below) cover FrontierMath, BullshitBench v2, Next.js Evals, BrowseComp, Humanity's Last Exam, and GDPval (win/tie rate).
Get OpenAI releases and AI news.
One email, every other Monday.
API pricing
- InputWhat you pay for everything you send the model — your question, plus any documents or earlier conversation you include with it.
- $30.00
- OutputWhat you pay for the text the model writes back. It is normally the dearer half: producing an answer costs more than reading one.
- $180.00
Benchmarks
Coding
Next.js EvalsNext.js coding — Vercel's open eval of how well AI coding agents build and migrate real Next.js apps — measured as the share of tasks the agent completes successfully. Higher is better.
65%
#16 of 28Best: Claude Fable 5.1 · 97%
Agentic & tool use
BrowseCompWeb browsing — Can the AI browse the web and track down hard-to-find answers? Higher is better.
90.1%
#5 of 28Best: GPT-5.6 Sol · 92.2%
Reasoning & science
FrontierMathAdvanced math — Very hard, research-level math problems. Tiers 1–3 are the (still extremely difficult) lower tiers. Higher is better.
52.4%
Tier 1–3
39.6%
Tier 4
Tier 1–3
#1Best published FrontierMath score of all tracked models
Tier 4
#1Best published FrontierMath score of all tracked models
Humanity's Last ExamMultidisciplinary reasoning — Humanity's Last Exam — extremely hard expert questions across many subjects. “With tools” means the AI is allowed to search the web or run code while answering. Higher is better.
#12 of 42Best: Claude Fable 5.1 · 65%
Knowledge work
GDPval (win/tie rate)Knowledge work — How often the AI's work matches or beats a human expert's on real knowledge-work tasks. Higher is better.
82.3%
#3 of 6Best: GPT-5.5 · 84.9%
Robustness
BullshitBench v2Nonsense detection — Given a confidently-worded but nonsensical prompt, does the AI spot that it makes no sense and push back — instead of playing along and inventing an answer? The score is how often it clearly called out the nonsense. Higher is better.
36%
#42 of 77Best: Claude Opus 4.8 · 95%
Source: BenchLM, retrieved 11 July 2026. Every other score here is the figure the lab published at launch.
Compare GPT-5.5-Pro with
Suggested comparisons
GPT-5.5-ProvsGPT-6 AstraGPT-5.5-ProvsClaude Fable 5.1GPT-5.5-ProvsGemini 3.8 FlashGPT-5.5-ProvsMuse Spark 1.3GPT-5.5-ProvsGrok 4.6GPT-5.5-ProvsDeepSeek-V4.1-FlashGPT-5.5-ProvsMistral Medium 3.5GPT-5.5-ProvsKimi K3GPT-5.5-ProvsGLM-5.3-FlashGPT-5.5-ProvsQwen3.8-Max-0902GPT-5.5-ProvsNemotron 3.5 Lightning
Frequently asked questions
GPT-5.5-Pro was released by OpenAI on Thursday, Apr 23 2026.
All OpenAI releases
44 tracked2026
14 releasesGPT-6 Astra
Sep 3 2026
GPT-5.6-Cyber
Aug 10 2026
GPT-5.6 Sol
Jun 26 2026
GPT-5.6 Terra
Jun 26 2026
GPT-5.6 Luna
Jun 26 2026
GPT-5.5-Cyber
Jun 22 2026
GPT-5.5
Apr 23 2026
GPT-5.5-Pro
Apr 23 2026
GPT-5.4 mini
Mar 17 2026
GPT-5.4 nano
Mar 17 2026
GPT-5.4
Mar 5 2026
GPT-5.4-Pro
Mar 5 2026
GPT-5.3-Codex-Spark
Feb 12 2026
GPT-5.3-Codex
Feb 5 2026
2025
22 releasesGPT-5.2
Dec 11 2025
GPT-5.1-Codex-Max
Nov 19 2025
GPT-5.1
Nov 12 2025
GPT-5-Codex-Mini
Nov 7 2025
GPT-5-Codex
Sep 15 2025
GPT-5
Aug 7 2025
GPT-5 mini
Aug 7 2025
GPT-5 nano
Aug 7 2025
GPT-5 Pro
Aug 7 2025
gpt-oss-120b
Aug 5 2025
gpt-oss-20b
Aug 5 2025
o3-pro
Jun 10 2025
o3
Apr 16 2025
o4-mini
Apr 16 2025
o4-mini-high
Apr 16 2025
GPT-4.1
Apr 14 2025
GPT-4.1 mini
Apr 14 2025
GPT-4.1 nano
Apr 14 2025
o1-pro
Mar 19 2025
GPT-4.5
Feb 27 2025
o3-mini
Jan 31 2025
o3-mini-high
Jan 31 2025