Claude 3.5 SonnetvsGPT-5.6 Luna

Claude 3.5 Sonnet
GPT-5.6 Luna
Benchmarks
Nonsense detection
BullshitBench v2
45%Best
40%
Prompt injection robustness
Gray Swan IPI · k = 1
8.3%
Prompt injection robustness
Gray Swan IPI · k = 10
38.6%
Prompt injection robustness
Gray Swan IPI · k = 15
43.9%
Coding
SWE-Bench Verified
33.4%
Agentic coding
CursorBench v3.2
61.1%
Agentic computer work
Frontier-Bench v0.1
14.3%
Agentic terminal coding
Terminal-Bench 2.1
82.5%
Science
GPQA Diamond
59.4%
Community preference (code)
Arena Elo (Code)
1523
Overview
CompanyAnthropicOpenAI
Release dateJun 20 2024Jun 26 2026
AccessProprietaryProprietary

Which is better: Claude 3.5 Sonnet or GPT-5.6 Luna?

Claude 3.5 Sonnet leads GPT-5.6 Luna on 1 of the 1 benchmark they both report (BullshitBench v2). Claude 3.5 Sonnet shipped 736 days before GPT-5.6 Luna, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Claude 3.5 Sonnet leads at 45% vs GPT-5.6 Luna at 40%.

Frequently asked questions

Claude 3.5 Sonnet was released by Anthropic on Jun 20 2024.

GPT-5.6 Luna was released by OpenAI on Jun 26 2026.

Other comparisons