Claude Sonnet 5vsGPT-5.6 Sol

Claude Sonnet 5
GPT-5.6 Sol
Benchmarks
Nonsense detection
BullshitBench v2
80%Best
47%
Prompt injection robustness
Gray Swan IPI · k = 1
0.6%
3.1%Best
Prompt injection robustness
Gray Swan IPI · k = 10
4.7%
16.3%Best
Prompt injection robustness
Gray Swan IPI · k = 15
5.9%
20%Best
Agentic coding
SWE-Bench Pro
63.2%
Agentic coding
CursorBench v3.2
61.5%
67.2%Best
Agentic coding
CursorBench v3.1
61.2%
Agentic coding
DeepSWE 1.1
73%
Next.js coding
Next.js Evals
79%
92%Best
Supabase coding
Supabase Evals · with skills
94.7%
100%Best
Supabase coding
Supabase Evals · no skills
78.9%
94.7%Best
Agentic computer work
Frontier-Bench v0.1
14.6%
34.4%Best
Agentic terminal coding
Terminal-Bench 2.1
80.4%
88.8%Best
Browser agent
BU Bench
67%
Web browsing
BrowseComp
84.7%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
43.2%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
57.4%
Agentic computer use
OSWorld-Verified
81.2%
Knowledge work
GDPval-AA
1618
Knowledge work
GDPval-AA v2
1748
Community preference
Arena Elo (Text)
1486
Community preference (code)
Arena Elo (Code)
1543
1620Best
Overview
CompanyAnthropicOpenAI
Release dateJun 30 2026Jun 26 2026
AccessProprietaryProprietary

Which is better: Claude Sonnet 5 or GPT-5.6 Sol?

GPT-5.6 Sol leads Claude Sonnet 5 on 7 of the 11 benchmarks they both report. GPT-5.6 Sol shipped 4 days before Claude Sonnet 5, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Claude Sonnet 5 leads at 80% vs GPT-5.6 Sol at 47%. On Gray Swan IPI · k = 1, Claude Sonnet 5 leads at 0.6% vs GPT-5.6 Sol at 3.1%. On Gray Swan IPI · k = 10, Claude Sonnet 5 leads at 4.7% vs GPT-5.6 Sol at 16.3%. On Gray Swan IPI · k = 15, Claude Sonnet 5 leads at 5.9% vs GPT-5.6 Sol at 20%. On CursorBench v3.2, GPT-5.6 Sol leads at 67.2% vs Claude Sonnet 5 at 61.5%. On Next.js Evals, GPT-5.6 Sol leads at 92% vs Claude Sonnet 5 at 79%. On Supabase Evals · with skills, GPT-5.6 Sol leads at 100% vs Claude Sonnet 5 at 94.7%. On Supabase Evals · no skills, GPT-5.6 Sol leads at 94.7% vs Claude Sonnet 5 at 78.9%. On Frontier-Bench v0.1, GPT-5.6 Sol leads at 34.4% vs Claude Sonnet 5 at 14.6%. On Terminal-Bench 2.1, GPT-5.6 Sol leads at 88.8% vs Claude Sonnet 5 at 80.4%. On Arena Elo (Code), GPT-5.6 Sol leads at 1620 vs Claude Sonnet 5 at 1543.

Frequently asked questions

Claude Sonnet 5 was released by Anthropic on Jun 30 2026.

GPT-5.6 Sol was released by OpenAI on Jun 26 2026.

GPT-5.6 Sol leads on CursorBench v3.2 — Claude Sonnet 5 61.5% vs GPT-5.6 Sol 67.2%.

Other comparisons