DeepSeek-V4.1-FlashvsQwen3.8-Max

DeepSeek-V4.1-Flash
Qwen3.8-Max
Specifications
Parameters
2.4T
Context window
1M
API pricing
Cheapest input
$2.00Alibaba
Cheapest output
$6.00Alibaba
Benchmarks
DeepSWE 1.1
74.2%56.6%
NL2Repo-Bench
65.4%55.9%
Terminal-Bench 3.0
30%11.3%
Terminal-Bench 2.1
90.6%86.6%
Humanity's Last Exam · with tools
63.9%43.6%
Benchmarks
BullshitBench v2
94%
SWE-Bench Pro
67.7%
PaperBench
93%
QwenSWEBench V2
55.1%
Terminal-Bench 4.0
31.2%
JobBench
53.4%
CoWorkBench
74.8%
Toolathlon-Verified
72.5%
CyberGym
88.1%
Humanity's Last Exam · no tools
36.8%
GPQA Diamond
90.9%
OSWorld-Verified
86.1%
Agent's Last Exam · pass@1
31.8%
AutomationBench
54.8%
CharXiv Reasoning
88.4%
Chartography · with tools
78.9%
BabyVision
82%
MMMU-Pro
82.3%
Overview
CompanyDeepSeekQwen
Release dateSep 10 2026Aug 3 2026
AccessProprietaryProprietary

Other comparisons

DeepSeek-V4.1-FlashvsClaude Fable 5.1Qwen3.8-MaxvsClaude Fable 5.1DeepSeek-V4.1-FlashvsGPT-6 AstraQwen3.8-MaxvsGPT-6 AstraDeepSeek-V4.1-FlashvsGemini 3.8 FlashQwen3.8-MaxvsGemini 3.8 FlashDeepSeek-V4.1-FlashvsMuse Spark 1.3Qwen3.8-MaxvsMuse Spark 1.3DeepSeek-V4.1-FlashvsGrok 4.6Qwen3.8-MaxvsGrok 4.6DeepSeek-V4.1-FlashvsMistral Medium 3.5Qwen3.8-MaxvsMistral Medium 3.5

Frequently asked questions

DeepSeek-V4.1-Flash leads Qwen3.8-Max on 5 of the 5 benchmarks they both report. Qwen3.8-Max shipped 38 days before DeepSeek-V4.1-Flash, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On DeepSWE 1.1, DeepSeek-V4.1-Flash leads at 74.2% vs Qwen3.8-Max at 56.6%. On NL2Repo-Bench, DeepSeek-V4.1-Flash leads at 65.4% vs Qwen3.8-Max at 55.9%. On Terminal-Bench 3.0, DeepSeek-V4.1-Flash leads at 30% vs Qwen3.8-Max at 11.3%. On Terminal-Bench 2.1, DeepSeek-V4.1-Flash leads at 90.6% vs Qwen3.8-Max at 86.6%. On Humanity's Last Exam · with tools, DeepSeek-V4.1-Flash leads at 63.9% vs Qwen3.8-Max at 43.6%.