DeepSeek-V4.1-FlashvsGPT-5.5-Cyber

DeepSeek-V4.1-Flash
GPT-5.5-Cyber
Benchmarks
CyberGym
88.1%85.6%
Benchmarks
DeepSWE 1.1
74.2%
NL2Repo-Bench
65.4%
Terminal-Bench 4.0
31.2%
Terminal-Bench 3.0
30%
Terminal-Bench 2.1
90.6%
Humanity's Last Exam · no tools
36.8%
Humanity's Last Exam · with tools
63.9%
GPQA Diamond
90.9%
Agent's Last Exam · pass@1
31.8%
AutomationBench
54.8%
Chartography · with tools
78.9%
Overview
CompanyDeepSeekOpenAI
Release dateSep 10 2026Jun 22 2026
AccessProprietaryProprietary

Other comparisons

DeepSeek-V4.1-FlashvsClaude Fable 5.1GPT-5.5-CybervsClaude Fable 5.1DeepSeek-V4.1-FlashvsGemini 3.8 FlashGPT-5.5-CybervsGemini 3.8 FlashDeepSeek-V4.1-FlashvsMuse Spark 1.3GPT-5.5-CybervsMuse Spark 1.3DeepSeek-V4.1-FlashvsGrok 4.6GPT-5.5-CybervsGrok 4.6DeepSeek-V4.1-FlashvsMistral Medium 3.5GPT-5.5-CybervsMistral Medium 3.5DeepSeek-V4.1-FlashvsKimi K3GPT-5.5-CybervsKimi K3

Frequently asked questions

DeepSeek-V4.1-Flash leads GPT-5.5-Cyber on 1 of the 1 benchmark they both report (CyberGym). GPT-5.5-Cyber shipped 80 days before DeepSeek-V4.1-Flash, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On CyberGym, DeepSeek-V4.1-Flash leads at 88.1% vs GPT-5.5-Cyber at 85.6%.