AI Model Release Tracker - Analytics

Claude 3.7 SonnetvsDeepSeek-V4-Pro

Claude 3.7 Sonnet
DeepSeek-V4-Pro
Benchmarks
Nonsense detection
BullshitBench v2
49%Best
14%
Coding
SWE-Bench Verified
62.3%
Competitive coding
LiveCodeBench
93.5%
Web browsing
BrowseComp
83.4%
Science
GPQA Diamond
68%
90.1%Best
Community preference (code)
Arena Elo (Code)
1457
Timeline
Release gapClaude 3.7 Sonnet shipped 424 days before DeepSeek-V4-Pro
Overview
CompanyAnthropicDeepSeek
Release dateFeb 24 2025Apr 24 2026
AccessProprietaryOpen Weight

Which is better: Claude 3.7 Sonnet or DeepSeek-V4-Pro?

Claude 3.7 Sonnet and DeepSeek-V4-Pro are evenly matched across the 2 benchmarks they both report (BullshitBench v2, GPQA Diamond). Claude 3.7 Sonnet shipped 424 days before DeepSeek-V4-Pro, so benchmark comparisons should account for the intervening progress.

Claude 3.7 Sonnet is proprietary, while DeepSeek-V4-Pro is open weight.

On BullshitBench v2, Claude 3.7 Sonnet leads at 49% vs DeepSeek-V4-Pro at 14%. On GPQA Diamond, DeepSeek-V4-Pro leads at 90.1% vs Claude 3.7 Sonnet at 68%.

Frequently asked questions

Claude 3.7 Sonnet was released by Anthropic on Feb 24 2025.

DeepSeek-V4-Pro was released by DeepSeek on Apr 24 2026.

DeepSeek-V4-Pro leads on GPQA Diamond — Claude 3.7 Sonnet 68% vs DeepSeek-V4-Pro 90.1%.

Claude 3.7 Sonnet is a proprietary model released by Anthropic. DeepSeek-V4-Pro is an open weight model released by DeepSeek.

Other comparisons