Claude Opus 4.7vsDeepSeek-V4.1-Flash

Claude Opus 4.7
DeepSeek-V4.1-Flash
Specifications
Context window
1M
API pricing
Input price
$5.00
Output price
$25.00
Cached input price
$0.50
Cheapest input
$5.00Amazon Bedrock
Cheapest output
$25.00Amazon Bedrock
Benchmarks
Terminal-Bench 2.1
66.1%90.6%
CyberGym
73.1%88.1%
Humanity's Last Exam · no tools
46.9%36.8%
Humanity's Last Exam · with tools
54.7%63.9%
GPQA Diamond
94.2%90.9%
Benchmarks
BullshitBench v2
83%
SWE-Bench Pro
64.3%
SWE-Bench Verified
87.6%
SWE-Bench Multilingual
80.5%
DeepSWE 1.1
74.2%
Next.js Evals
69%
NL2Repo-Bench
65.4%
Terminal-Bench 4.0
31.2%
Terminal-Bench 3.0
30%
Terminal-Bench 2.0
69.4%
MCP Atlas
79.1%
BrowseComp
79.3%
ARC-AGI-2
75.8%
FrontierMath · Tier 1–3
43.8%
FrontierMath · Tier 4
22.9%
OSWorld-Verified
78%
Agent's Last Exam · pass@1
31.8%
AutomationBench
54.8%
Finance Agent v2
51.5%
GDPval-AA
1753
GDPval (win/tie rate)
80.3%
CharXiv Reasoning
82.1%
Chartography · with tools
78.9%
MMMU-Pro
75.2%
Blueprint-Bench 2
24.5%
MRCR v2 (8-needle) · 128k average
59.3%
Overview
CompanyAnthropicDeepSeek
Release dateApr 16 2026Sep 10 2026
AccessProprietaryProprietary

Other comparisons

Claude Opus 4.7vsGPT-6 AstraDeepSeek-V4.1-FlashvsGPT-6 AstraClaude Opus 4.7vsGemini 3.8 FlashDeepSeek-V4.1-FlashvsGemini 3.8 FlashClaude Opus 4.7vsMuse Spark 1.3DeepSeek-V4.1-FlashvsMuse Spark 1.3Claude Opus 4.7vsGrok 4.6DeepSeek-V4.1-FlashvsGrok 4.6Claude Opus 4.7vsMistral Medium 3.5DeepSeek-V4.1-FlashvsMistral Medium 3.5Claude Opus 4.7vsKimi K3DeepSeek-V4.1-FlashvsKimi K3

Frequently asked questions

DeepSeek-V4.1-Flash leads Claude Opus 4.7 on 3 of the 5 benchmarks they both report (Terminal-Bench 2.1, CyberGym, Humanity's Last Exam, GPQA Diamond). Only Claude Opus 4.7 has a verified first-party API price: $5.00 per million input tokens and $25.00 per million output tokens. No pay-as-you-go API rate is tracked for DeepSeek-V4.1-Flash. Claude Opus 4.7 shipped 147 days before DeepSeek-V4.1-Flash, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On Terminal-Bench 2.1, DeepSeek-V4.1-Flash leads at 90.6% vs Claude Opus 4.7 at 66.1%. On CyberGym, DeepSeek-V4.1-Flash leads at 88.1% vs Claude Opus 4.7 at 73.1%. On Humanity's Last Exam · no tools, Claude Opus 4.7 leads at 46.9% vs DeepSeek-V4.1-Flash at 36.8%. On Humanity's Last Exam · with tools, DeepSeek-V4.1-Flash leads at 63.9% vs Claude Opus 4.7 at 54.7%. On GPQA Diamond, Claude Opus 4.7 leads at 94.2% vs DeepSeek-V4.1-Flash at 90.9%.