AI Model Release Tracker - Analytics

Claude Opus 4.1vsDeepSeek-V4-Pro

Claude Opus 4.1
DeepSeek-V4-Pro
Benchmarks
Nonsense detection
BullshitBench v2
43%Best
14%
Coding
SWE-Bench Verified
74.5%
Competitive coding
LiveCodeBench
93.5%
Web browsing
BrowseComp
83.4%
Science
GPQA Diamond
80.9%
90.1%Best
Community preference (code)
Arena Elo (Code)
1457
Overview
CompanyAnthropicDeepSeek
Release dateAug 5 2025Apr 24 2026
AccessProprietaryOpen Weight

Which is better: Claude Opus 4.1 or DeepSeek-V4-Pro?

Claude Opus 4.1 and DeepSeek-V4-Pro are evenly matched across the 2 benchmarks they both report (BullshitBench v2, GPQA Diamond). Claude Opus 4.1 shipped 262 days before DeepSeek-V4-Pro, so benchmark comparisons should account for the intervening progress.

Claude Opus 4.1 is proprietary, while DeepSeek-V4-Pro is open weight.

On BullshitBench v2, Claude Opus 4.1 leads at 43% vs DeepSeek-V4-Pro at 14%. On GPQA Diamond, DeepSeek-V4-Pro leads at 90.1% vs Claude Opus 4.1 at 80.9%.

Frequently asked questions

Claude Opus 4.1 was released by Anthropic on Aug 5 2025.

DeepSeek-V4-Pro was released by DeepSeek on Apr 24 2026.

DeepSeek-V4-Pro leads on GPQA Diamond — Claude Opus 4.1 80.9% vs DeepSeek-V4-Pro 90.1%.

Claude Opus 4.1 is a proprietary model released by Anthropic. DeepSeek-V4-Pro is an open weight model released by DeepSeek.

Other comparisons