DeepSeek-V4-Flash-0731vsQwen3.8-Max

DeepSeek-V4-Flash-0731
Qwen3.8-Max
Specifications
Parameters
2.4T
Context window
1M
API pricing
Input price
$0.22
Output price
$0.66
Cached input price
$0.007
Cheapest input
$0.03OpenInference$2.00Alibaba
Cheapest output
$0.07OpenInference$6.00Alibaba
Benchmarks
BullshitBench v2
39%95%
Terminal-Bench 2.1
82.7%86.6%
Toolathlon-Verified
70.3%72.5%
Humanity's Last Exam · with tools
34.8%43.6%
Benchmarks
SWE-Bench Pro
67.7%
SWE-Bench Verified
79%
DeepSWE 1.1
56.6%
PaperBench
93%
NL2Repo-Bench
55.9%
QwenSWEBench V2
55.1%
Terminal-Bench 3.0
11.3%
JobBench
53.4%
CoWorkBench
74.8%
BrowseComp
73.2%
CyberGym
76.7%
OSWorld-Verified
86.1%
AutomationBench
25.1%
CharXiv Reasoning
88.4%
BabyVision
82%
MMMU-Pro
82.3%
Overview
CompanyDeepSeekQwen
Release dateJul 31 2026Aug 3 2026
AccessProprietaryProprietary

Other comparisons

DeepSeek-V4-Flash-0731vsClaude Fable 5.1Qwen3.8-MaxvsClaude Fable 5.1DeepSeek-V4-Flash-0731vsGPT-6 AstraQwen3.8-MaxvsGPT-6 AstraDeepSeek-V4-Flash-0731vsGemini 3.8 FlashQwen3.8-MaxvsGemini 3.8 FlashDeepSeek-V4-Flash-0731vsMuse Spark 1.3Qwen3.8-MaxvsMuse Spark 1.3DeepSeek-V4-Flash-0731vsGrok 4.6Qwen3.8-MaxvsGrok 4.6DeepSeek-V4-Flash-0731vsMistral Medium 3.5Qwen3.8-MaxvsMistral Medium 3.5

Frequently asked questions

Qwen3.8-Max leads DeepSeek-V4-Flash-0731 on 4 of the 4 benchmarks they both report (BullshitBench v2, Terminal-Bench 2.1, Toolathlon-Verified, Humanity's Last Exam). Only DeepSeek-V4-Flash-0731 has a verified first-party API price: $0.22 per million input tokens and $0.66 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Max. DeepSeek-V4-Flash-0731 shipped 3 days before Qwen3.8-Max, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Qwen3.8-Max leads at 95% vs DeepSeek-V4-Flash-0731 at 39%. On Terminal-Bench 2.1, Qwen3.8-Max leads at 86.6% vs DeepSeek-V4-Flash-0731 at 82.7%. On Toolathlon-Verified, Qwen3.8-Max leads at 72.5% vs DeepSeek-V4-Flash-0731 at 70.3%. On Humanity's Last Exam · with tools, Qwen3.8-Max leads at 43.6% vs DeepSeek-V4-Flash-0731 at 34.8%.