AI Model Release Tracker - Analytics

Claude Opus 4.6vsDeepSeek-V4-Pro

Claude Opus 4.6
DeepSeek-V4-Pro
Benchmarks
Nonsense detection
BullshitBench v2
87%Best
14%
Coding
SWE-Bench Verified
80.8%
Next.js coding
Next.js Evals
75%
Competitive coding
LiveCodeBench
93.5%
Web browsing
BrowseComp
83.7%Best
83.4%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
53%
Science
GPQA Diamond
91.3%Best
90.1%
Community preference
Arena Elo (Text)
1504
Community preference (code)
Arena Elo (Code)
1543Best
1457
Timeline
Release gapClaude Opus 4.6 shipped 78 days before DeepSeek-V4-Pro
Overview
CompanyAnthropicDeepSeek
Release dateFeb 5 2026Apr 24 2026
AccessProprietaryOpen Weight

Which is better: Claude Opus 4.6 or DeepSeek-V4-Pro?

Claude Opus 4.6 leads DeepSeek-V4-Pro on 4 of the 4 benchmarks they both report (BullshitBench v2, BrowseComp, GPQA Diamond, Arena Elo (Code)). Claude Opus 4.6 shipped 78 days before DeepSeek-V4-Pro, so benchmark comparisons should account for the intervening progress.

Claude Opus 4.6 is proprietary, while DeepSeek-V4-Pro is open weight.

On BullshitBench v2, Claude Opus 4.6 leads at 87% vs DeepSeek-V4-Pro at 14%. On BrowseComp, Claude Opus 4.6 leads at 83.7% vs DeepSeek-V4-Pro at 83.4%. On GPQA Diamond, Claude Opus 4.6 leads at 91.3% vs DeepSeek-V4-Pro at 90.1%. On Arena Elo (Code), Claude Opus 4.6 leads at 1543 vs DeepSeek-V4-Pro at 1457.

Frequently asked questions

Claude Opus 4.6 was released by Anthropic on Feb 5 2026.

DeepSeek-V4-Pro was released by DeepSeek on Apr 24 2026.

Claude Opus 4.6 leads on GPQA Diamond — Claude Opus 4.6 91.3% vs DeepSeek-V4-Pro 90.1%.

Claude Opus 4.6 is a proprietary model released by Anthropic. DeepSeek-V4-Pro is an open weight model released by DeepSeek.

Other comparisons