GPT-5.5vsQwen3.8-Flash-Next

GPT-5.5
Qwen3.8-Flash-Next
Specifications
Parameters
125B
Context window
1.05M262k
API pricing
Input price
$5.00
Output price
$30.00
Cached input price
$0.50
Cheapest input
$5.00Azure
Cheapest output
$30.00Azure
Benchmarks
SWE-Bench Pro
58.6%62.5%
SWE-Bench Multilingual
77.8%81%
Humanity's Last Exam · no tools
41.4%35.9%
GPQA Diamond
93.6%91.7%
CharXiv Reasoning
84.1%84.6%
Benchmarks
BullshitBench v2
47%
Gray Swan IPI · k = 1
3%
Gray Swan IPI · k = 10
17.4%
Gray Swan IPI · k = 15
20.8%
DeepSWE 1.1
58.7%
DeepSWE 1.0
64.3%
NL2Repo-Bench
48.1%
LiveCodeBench
91.9%
Terminal-Bench 2.1
78.2%
Terminal-Bench 2.0
82.7%
Expert-SWE (Internal)
73.1%
MCP Atlas
75.3%
JobBench
55.7%
CoWorkBench
73.9%
Toolathlon-Verified
73.5%
Toolathlon
55.6%
BrowseComp
84.4%
CyberGym
81.8%
Humanity's Last Exam · with tools
52.2%
ARC-AGI-2
84.6%
FrontierMath · Tier 1–3
51.7%
FrontierMath · Tier 4
35.4%
IFBench
81.3%
OSWorld 2.0
19.4%
OSWorld-Verified
78.7%
Agent's Last Exam · pass@1
24.3%
Agent's Last Exam · score
51.2%
Finance Agent v2
51.8%
Harvey's Legal Agent Benchmark
3.75%
TaxEval v2
74.98%
MedScribe
86.87%
GDPval-AA
1769
GDPval-AA v2
1494
GDPval (win/tie rate)
84.9%
LVBench
76.6%
MMMU-Pro
81.2%
Blueprint-Bench 2
36.2%
MRCR v2 (8-needle) · 128k average
94.8%
Overview
CompanyOpenAIQwen
Release dateApr 23 2026Aug 26 2026
AccessProprietaryOpen Weight

Other comparisons

GPT-5.5vsClaude Opus 5Qwen3.8-Flash-NextvsClaude Opus 5GPT-5.5vsGemini 3.7 FlashQwen3.8-Flash-NextvsGemini 3.7 FlashGPT-5.5vsMuse GlimmerQwen3.8-Flash-NextvsMuse GlimmerGPT-5.5vsGrok 4.6Qwen3.8-Flash-NextvsGrok 4.6GPT-5.5vsDeepSeek-V4-Pro-0813Qwen3.8-Flash-NextvsDeepSeek-V4-Pro-0813GPT-5.5vsMistral Medium 3.5Qwen3.8-Flash-NextvsMistral Medium 3.5

Frequently asked questions

Qwen3.8-Flash-Next leads GPT-5.5 on 3 of the 5 benchmarks they both report. Only GPT-5.5 has a verified first-party API price: $5.00 per million input tokens and $30.00 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. GPT-5.5 shipped 125 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

Context windows are 1.05M (GPT-5.5) vs 262k (Qwen3.8-Flash-Next). GPT-5.5 is proprietary, while Qwen3.8-Flash-Next is open weight.

On SWE-Bench Pro, Qwen3.8-Flash-Next leads at 62.5% vs GPT-5.5 at 58.6%. On SWE-Bench Multilingual, Qwen3.8-Flash-Next leads at 81% vs GPT-5.5 at 77.8%. On Humanity's Last Exam · no tools, GPT-5.5 leads at 41.4% vs Qwen3.8-Flash-Next at 35.9%. On GPQA Diamond, GPT-5.5 leads at 93.6% vs Qwen3.8-Flash-Next at 91.7%. On CharXiv Reasoning, Qwen3.8-Flash-Next leads at 84.6% vs GPT-5.5 at 84.1%.

GPT-5.5 was released by OpenAI on Apr 23 2026.

Qwen3.8-Flash-Next was released by Qwen on Aug 26 2026.

Qwen3.8-Flash-Next leads on SWE-Bench Pro — GPT-5.5 58.6% vs Qwen3.8-Flash-Next 62.5%.

GPT-5.5 leads on Humanity's Last Exam · no tools — GPT-5.5 41.4% vs Qwen3.8-Flash-Next 35.9%.

Only GPT-5.5 has a verified first-party API price: $5.00 per million input tokens and $30.00 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. Rates are pay-as-you-go API prices verified on August 18, 2026.

GPT-5.5 has a 1.05M context window; Qwen3.8-Flash-Next has 262k.

GPT-5.5 is a proprietary model released by OpenAI. Qwen3.8-Flash-Next is an open weight model released by Qwen.