DeepSeek-V4.1-FlashvsGPT-6 Astra

DeepSeek-V4.1-Flash
GPT-6 Astra
Specifications
Context window
1.05M
API pricing
Input price
$10.00
Output price
$50.00
Cached input price
$1.00
Cheapest input
$11.00Azure
Cheapest output
$55.00Azure
Benchmarks
DeepSWE 1.1
74.2%74.1%
Terminal-Bench 4.0
31.2%57.9%
GPQA Diamond
90.9%96%
Agent's Last Exam · pass@1
31.8%59.3%
AutomationBench
54.8%41.4%
Benchmarks
Auto-review circumvention (Internal)
0%
FrontierCode v1.1 (Main) · main split
53.3%
FrontierCode v1.1 (Extended) · extended split
64.5%
Next.js Evals
85%
NL2Repo-Bench
65.4%
AA Coding Agent Index
67
Database Migration Tasks (OpenAI Internal)
63.9%
BenchCAD
95.9%
Terminal-Bench 3.0
30%
Terminal-Bench 2.1
90.6%
Terminal-Bench-Science 0.1
64.6%
BrowseComp
91.5%
CyberGym
88.1%
ExploitBench
100%
Humanity's Last Exam · no tools
36.8%
Humanity's Last Exam · with tools
63.9%
ARC-AGI-3
99.9%
ARC-AGI-2
95%
FrontierMath · Tier 4 (v2)
97.6%
GeneBench-Pro
39%
MedChemBench (Internal)
49.7%
SRE-Bench
99.2%
HealthBench Professional · length-adjusted
63.4%
AA Intelligence Index
61.2
Design Tasks (OpenAI Internal)
50%
Data Science Tasks (OpenAI Internal)
40.9%
Chartography · with tools
78.9%
OpenScore String Quartets
0.84
threejseval
2001
Overview
CompanyDeepSeekOpenAI
Release dateSep 10 2026Sep 3 2026
AccessProprietaryProprietary

Other comparisons

DeepSeek-V4.1-FlashvsClaude Fable 5.1GPT-6 AstravsClaude Fable 5.1DeepSeek-V4.1-FlashvsGemini 3.8 FlashGPT-6 AstravsGemini 3.8 FlashDeepSeek-V4.1-FlashvsMuse Spark 1.3GPT-6 AstravsMuse Spark 1.3DeepSeek-V4.1-FlashvsGrok 4.6GPT-6 AstravsGrok 4.6DeepSeek-V4.1-FlashvsMistral Medium 3.5GPT-6 AstravsMistral Medium 3.5DeepSeek-V4.1-FlashvsKimi K3GPT-6 AstravsKimi K3

Frequently asked questions

GPT-6 Astra leads DeepSeek-V4.1-Flash on 3 of the 5 benchmarks they both report. Only GPT-6 Astra has a verified first-party API price: $10.00 per million input tokens and $50.00 per million output tokens. No pay-as-you-go API rate is tracked for DeepSeek-V4.1-Flash. GPT-6 Astra shipped 7 days before DeepSeek-V4.1-Flash, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On DeepSWE 1.1, DeepSeek-V4.1-Flash leads at 74.2% vs GPT-6 Astra at 74.1%. On Terminal-Bench 4.0, GPT-6 Astra leads at 57.9% vs DeepSeek-V4.1-Flash at 31.2%. On GPQA Diamond, GPT-6 Astra leads at 96% vs DeepSeek-V4.1-Flash at 90.9%. On Agent's Last Exam · pass@1, GPT-6 Astra leads at 59.3% vs DeepSeek-V4.1-Flash at 31.8%. On AutomationBench, DeepSeek-V4.1-Flash leads at 54.8% vs GPT-6 Astra at 41.4%.