Claude Fable 5.1vsDeepSeek-V4.1-Flash

Claude Fable 5.1
DeepSeek-V4.1-Flash
Specifications
Context window
1M
API pricing
Input price
$10.00
Output price
$50.00
Cached input price
$0.25
Cheapest input
$10.00Amazon Bedrock
Cheapest output
$50.00Amazon Bedrock
Benchmarks
Terminal-Bench 4.0
57.88%31.2%
Humanity's Last Exam · no tools
60.9%36.8%
Humanity's Last Exam · with tools
65%63.9%
AutomationBench
31.4%54.8%
Benchmarks
SWE-Bench Pro
81.2%
SWE-Bench Multilingual
89.1%
SWE-Bench Multimodal
54.7%
DeepSWE 1.1
74.2%
NL2Repo-Bench
65.4%
Terminal-Bench 3.0
30%
Terminal-Bench 2.1
90.6%
Terminal-Bench-Science 0.1
52.6%
CyberGym
88.1%
ARC-AGI-2
90%
GPQA Diamond
90.9%
Agent's Last Exam · pass@1
31.8%
HealthBench Professional
62.1%
GDPval-AA v2
1853
AA-Briefcase
1694
Chartography · with tools
78.9%
threejseval
2037
Overview
CompanyAnthropicDeepSeek
Release dateSep 1 2026Sep 10 2026
AccessProprietaryProprietary

Other comparisons

Claude Fable 5.1vsGPT-6 AstraDeepSeek-V4.1-FlashvsGPT-6 AstraClaude Fable 5.1vsGemini 3.8 FlashDeepSeek-V4.1-FlashvsGemini 3.8 FlashClaude Fable 5.1vsMuse Spark 1.3DeepSeek-V4.1-FlashvsMuse Spark 1.3Claude Fable 5.1vsGrok 4.6DeepSeek-V4.1-FlashvsGrok 4.6Claude Fable 5.1vsMistral Medium 3.5DeepSeek-V4.1-FlashvsMistral Medium 3.5Claude Fable 5.1vsKimi K3DeepSeek-V4.1-FlashvsKimi K3

Frequently asked questions

Claude Fable 5.1 leads DeepSeek-V4.1-Flash on 3 of the 4 benchmarks they both report (Terminal-Bench 4.0, Humanity's Last Exam, AutomationBench). Only Claude Fable 5.1 has a verified first-party API price: $10.00 per million input tokens and $50.00 per million output tokens. No pay-as-you-go API rate is tracked for DeepSeek-V4.1-Flash. Claude Fable 5.1 shipped 9 days before DeepSeek-V4.1-Flash, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On Terminal-Bench 4.0, Claude Fable 5.1 leads at 57.88% vs DeepSeek-V4.1-Flash at 31.2%. On Humanity's Last Exam · no tools, Claude Fable 5.1 leads at 60.9% vs DeepSeek-V4.1-Flash at 36.8%. On Humanity's Last Exam · with tools, Claude Fable 5.1 leads at 65% vs DeepSeek-V4.1-Flash at 63.9%. On AutomationBench, DeepSeek-V4.1-Flash leads at 54.8% vs Claude Fable 5.1 at 31.4%.