Claude 3.7 SonnetvsGPT-5-Codex

Claude 3.7 Sonnet
GPT-5-Codex
Benchmarks
Nonsense detection
BullshitBench v2
49%Best
39%
Coding
SWE-Bench Verified
62.3%
Science
GPQA Diamond
68%
Overview
CompanyAnthropicOpenAI
Release dateFeb 24 2025Sep 15 2025
AccessProprietaryProprietary

Which is better: Claude 3.7 Sonnet or GPT-5-Codex?

Claude 3.7 Sonnet leads GPT-5-Codex on 1 of the 1 benchmark they both report (BullshitBench v2). Claude 3.7 Sonnet shipped 203 days before GPT-5-Codex, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Claude 3.7 Sonnet leads at 49% vs GPT-5-Codex at 39%.

Frequently asked questions

Claude 3.7 Sonnet was released by Anthropic on Feb 24 2025.

GPT-5-Codex was released by OpenAI on Sep 15 2025.

Other comparisons