DeepSeek-V4.1-FlashvsQwen3.8-Max-0902

DeepSeek-V4.1-Flash
Qwen3.8-Max-0902
Specifications
Parameters
2.4T
Context window
1M
Benchmarks
DeepSWE 1.1
74.2%69.3%
NL2Repo-Bench
65.4%64.9%
Terminal-Bench 3.0
30%29%
AutomationBench
54.8%50.8%
Benchmarks
QwenSWEBench V2
70%
Terminal-Bench 4.0
31.2%
Terminal-Bench 2.1
90.6%
JobBench
64%
CoWorkBench
76.1%
Toolathlon-Verified
73.3%
CyberGym
88.1%
Humanity's Last Exam · no tools
36.8%
Humanity's Last Exam · with tools
63.9%
GPQA Diamond
90.9%
Agent's Last Exam · pass@1
31.8%
Chartography · with tools
78.9%
MMMU-Pro
82.7%
Overview
CompanyDeepSeekQwen
Release dateSep 10 2026Sep 2 2026
AccessProprietaryProprietary

Other comparisons

DeepSeek-V4.1-FlashvsClaude Fable 5.1Qwen3.8-Max-0902vsClaude Fable 5.1DeepSeek-V4.1-FlashvsGPT-6 AstraQwen3.8-Max-0902vsGPT-6 AstraDeepSeek-V4.1-FlashvsGemini 3.8 FlashQwen3.8-Max-0902vsGemini 3.8 FlashDeepSeek-V4.1-FlashvsMuse Spark 1.3Qwen3.8-Max-0902vsMuse Spark 1.3DeepSeek-V4.1-FlashvsGrok 4.6Qwen3.8-Max-0902vsGrok 4.6DeepSeek-V4.1-FlashvsMistral Medium 3.5Qwen3.8-Max-0902vsMistral Medium 3.5

Frequently asked questions

DeepSeek-V4.1-Flash leads Qwen3.8-Max-0902 on 4 of the 4 benchmarks they both report (DeepSWE 1.1, NL2Repo-Bench, Terminal-Bench 3.0, AutomationBench). Qwen3.8-Max-0902 shipped 8 days before DeepSeek-V4.1-Flash, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On DeepSWE 1.1, DeepSeek-V4.1-Flash leads at 74.2% vs Qwen3.8-Max-0902 at 69.3%. On NL2Repo-Bench, DeepSeek-V4.1-Flash leads at 65.4% vs Qwen3.8-Max-0902 at 64.9%. On Terminal-Bench 3.0, DeepSeek-V4.1-Flash leads at 30% vs Qwen3.8-Max-0902 at 29%. On AutomationBench, DeepSeek-V4.1-Flash leads at 54.8% vs Qwen3.8-Max-0902 at 50.8%.