Claude Opus 4.7vsDeepSeek-V4-Flash-0731

Claude Opus 4.7
DeepSeek-V4-Flash-0731
Specifications
Context window
1M
API pricing
Input price
$5.00$0.22
Output price
$25.00$0.66
Cached input price
$0.50$0.007
Cheapest input
$5.00Amazon Bedrock$0.0352Baidu
Cheapest output
$25.00Amazon Bedrock$0.10OpenInference
Benchmarks
BullshitBench v2
83%39%
SWE-Bench Verified
87.6%79%
Terminal-Bench 2.1
66.1%82.7%
BrowseComp
79.3%73.2%
CyberGym
73.1%76.7%
Humanity's Last Exam · with tools
54.7%34.8%
Benchmarks
ProgramBench
0%
SWE-Bench Pro
64.3%
SWE-Bench Multilingual
80.5%
Next.js Evals
58%
Terminal-Bench 2.0
69.4%
MCP Atlas
79.1%
Toolathlon-Verified
70.3%
Humanity's Last Exam · no tools
46.9%
ARC-AGI-2
75.8%
FrontierMath · Tier 1–3
43.8%
FrontierMath · Tier 4
22.9%
GPQA Diamond
94.2%
OSWorld-Verified
78%
AutomationBench
25.1%
Finance Agent v2
51.5%
GDPval-AA
1753
GDPval (win/tie rate)
80.3%
CharXiv Reasoning
82.1%
MMMU-Pro
75.2%
Blueprint-Bench 2
24.5%
MRCR v2 (8-needle) · 128k average
59.3%
Overview
CompanyAnthropicDeepSeek
Release dateApr 16 2026Jul 31 2026
AccessProprietaryProprietary

Other comparisons

Claude Opus 4.7vsGPT-6 AstraDeepSeek-V4-Flash-0731vsGPT-6 AstraClaude Opus 4.7vsGemini 3.8 FlashDeepSeek-V4-Flash-0731vsGemini 3.8 FlashClaude Opus 4.7vsMuse Spark 1.3DeepSeek-V4-Flash-0731vsMuse Spark 1.3Claude Opus 4.7vsGrok 4.6DeepSeek-V4-Flash-0731vsGrok 4.6Claude Opus 4.7vsMistral Medium 3.5DeepSeek-V4-Flash-0731vsMistral Medium 3.5Claude Opus 4.7vsKimi K3DeepSeek-V4-Flash-0731vsKimi K3

Frequently asked questions

Claude Opus 4.7 leads DeepSeek-V4-Flash-0731 on 4 of the 6 benchmarks they both report. DeepSeek-V4-Flash-0731 is cheaper on both input and output: $0.22 vs $5.00 per million input tokens, and $0.66 vs $25.00 per million output tokens. Figures are base-tier rates. Claude Opus 4.7 shipped 106 days before DeepSeek-V4-Flash-0731, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Claude Opus 4.7 leads at 83% vs DeepSeek-V4-Flash-0731 at 39%. On SWE-Bench Verified, Claude Opus 4.7 leads at 87.6% vs DeepSeek-V4-Flash-0731 at 79%. On Terminal-Bench 2.1, DeepSeek-V4-Flash-0731 leads at 82.7% vs Claude Opus 4.7 at 66.1%. On BrowseComp, Claude Opus 4.7 leads at 79.3% vs DeepSeek-V4-Flash-0731 at 73.2%. On CyberGym, DeepSeek-V4-Flash-0731 leads at 76.7% vs Claude Opus 4.7 at 73.1%. On Humanity's Last Exam · with tools, Claude Opus 4.7 leads at 54.7% vs DeepSeek-V4-Flash-0731 at 34.8%.