Claude Sonnet 5vsDeepSeek-V4-Flash-0731

Claude Sonnet 5
DeepSeek-V4-Flash-0731
API pricing
Input price
$2.00$0.22
Output price
$10.00$0.66
Cached input price
$0.20$0.007
Cheapest input
$2.00Amazon Bedrock$0.0352Baidu
Cheapest output
$10.00Amazon Bedrock$0.10OpenInference
Benchmarks
BullshitBench v2
80%39%
SWE-Bench Verified
85.2%79%
Terminal-Bench 2.1
80.4%82.7%
BrowseComp
84.7%73.2%
Humanity's Last Exam · with tools
57.4%34.8%
Benchmarks
Gray Swan IPI · k = 1
0.6%
Gray Swan IPI · k = 10
4.7%
Gray Swan IPI · k = 15
5.9%
SWE-Bench Pro
63.2%
Next.js Evals
81%
Supabase Evals · with skills
91.3%
Supabase Evals · no skills
79.7%
Frontier-Bench v0.1
14.6%
Terminal-Bench 4.0
12.42%
Toolathlon-Verified
70.3%
CyberGym
76.7%
Humanity's Last Exam · no tools
43.2%
OSWorld-Verified
81.2%
AutomationBench
25.1%
GDPval-AA
1618
threejseval
1273
Overview
CompanyAnthropicDeepSeek
Release dateJun 30 2026Jul 31 2026
AccessProprietaryProprietary

Other comparisons

Claude Sonnet 5vsGPT-6 AstraDeepSeek-V4-Flash-0731vsGPT-6 AstraClaude Sonnet 5vsGemini 3.8 FlashDeepSeek-V4-Flash-0731vsGemini 3.8 FlashClaude Sonnet 5vsMuse Spark 1.3DeepSeek-V4-Flash-0731vsMuse Spark 1.3Claude Sonnet 5vsGrok 4.6DeepSeek-V4-Flash-0731vsGrok 4.6Claude Sonnet 5vsMistral Medium 3.5DeepSeek-V4-Flash-0731vsMistral Medium 3.5Claude Sonnet 5vsKimi K3DeepSeek-V4-Flash-0731vsKimi K3

Frequently asked questions

Claude Sonnet 5 leads DeepSeek-V4-Flash-0731 on 4 of the 5 benchmarks they both report. DeepSeek-V4-Flash-0731 is cheaper on both input and output: $0.22 vs $2.00 per million input tokens, and $0.66 vs $10.00 per million output tokens. Figures are base-tier rates. Claude Sonnet 5 shipped 31 days before DeepSeek-V4-Flash-0731, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Claude Sonnet 5 leads at 80% vs DeepSeek-V4-Flash-0731 at 39%. On SWE-Bench Verified, Claude Sonnet 5 leads at 85.2% vs DeepSeek-V4-Flash-0731 at 79%. On Terminal-Bench 2.1, DeepSeek-V4-Flash-0731 leads at 82.7% vs Claude Sonnet 5 at 80.4%. On BrowseComp, Claude Sonnet 5 leads at 84.7% vs DeepSeek-V4-Flash-0731 at 73.2%. On Humanity's Last Exam · with tools, Claude Sonnet 5 leads at 57.4% vs DeepSeek-V4-Flash-0731 at 34.8%.