Compare AI models

API pricing
Input price
$3.00$0.66
Output price
$15.00$1.98
Cached input price
$0.30$0.022
Cheapest input
$3.00Amazon Bedrock$0.2442Baidu
Cheapest output
$15.00Amazon Bedrock$0.7326Baidu
Benchmarks
BullshitBench v2
91%35%
SWE-Bench Verified
79.6%80.6%
Humanity's Last Exam · no tools
33.2%42.7%
Humanity's Last Exam · with tools
49%60%
ARC-AGI-2
58.3%61.3%
Overview
CompanyAnthropicDeepSeek
Release dateFeb 17 2026Aug 13 2026
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

DeepSeek-V4-Pro-0813 leads Claude Sonnet 4.6 on 4 of the 5 benchmarks they both report (BullshitBench v2, SWE-Bench Verified, Humanity's Last Exam, ARC-AGI-2). DeepSeek-V4-Pro-0813 is cheaper on both input and output: $0.66 vs $3.00 per million input tokens, and $1.98 vs $15.00 per million output tokens. Figures are base-tier rates. Claude Sonnet 4.6 shipped 177 days before DeepSeek-V4-Pro-0813, so benchmark comparisons should account for the intervening progress.

Claude Sonnet 4.6 is closed, while DeepSeek-V4-Pro-0813 is open weight.

On BullshitBench v2, Claude Sonnet 4.6 leads at 91% vs DeepSeek-V4-Pro-0813 at 35%. On SWE-Bench Verified, DeepSeek-V4-Pro-0813 leads at 80.6% vs Claude Sonnet 4.6 at 79.6%. On Humanity's Last Exam · no tools, DeepSeek-V4-Pro-0813 leads at 42.7% vs Claude Sonnet 4.6 at 33.2%. On Humanity's Last Exam · with tools, DeepSeek-V4-Pro-0813 leads at 60% vs Claude Sonnet 4.6 at 49%. On ARC-AGI-2, DeepSeek-V4-Pro-0813 leads at 61.3% vs Claude Sonnet 4.6 at 58.3%.