gpt-oss-120bvsQwen3.8-Flash-Next

gpt-oss-120b
Qwen3.8-Flash-Next
Specifications
Parameters
117B125B
Context window
128k262k
API pricing
Cheapest input
$0.03AkashML
Cheapest output
$0.17AkashML
Benchmarks
Humanity's Last Exam · no tools
14.9%35.9%
GPQA Diamond
80.1%91.7%
Benchmarks
BullshitBench v2
11%
SWE-Bench Pro
62.5%
SWE-Bench Verified
62.4%
SWE-Bench Multilingual
81%
DeepSWE 1.1
58.7%
NL2Repo-Bench
48.1%
LiveCodeBench
91.9%
JobBench
55.7%
CoWorkBench
73.9%
Toolathlon-Verified
73.5%
Humanity's Last Exam · with tools
19%
IFBench
81.3%
MMLU
90%
OSWorld 2.0
19.4%
Agent's Last Exam · pass@1
24.3%
Agent's Last Exam · score
51.2%
CharXiv Reasoning
84.6%
LVBench
76.6%
Overview
CompanyOpenAIQwen
Release dateAug 5 2025Aug 26 2026
AccessOpen WeightOpen Weight

Other comparisons

gpt-oss-120bvsClaude Opus 5Qwen3.8-Flash-NextvsClaude Opus 5gpt-oss-120bvsGemini 3.7 FlashQwen3.8-Flash-NextvsGemini 3.7 Flashgpt-oss-120bvsMuse GlimmerQwen3.8-Flash-NextvsMuse Glimmergpt-oss-120bvsGrok 4.6Qwen3.8-Flash-NextvsGrok 4.6gpt-oss-120bvsDeepSeek-V4-Pro-0813Qwen3.8-Flash-NextvsDeepSeek-V4-Pro-0813gpt-oss-120bvsMistral Medium 3.5Qwen3.8-Flash-NextvsMistral Medium 3.5

Frequently asked questions

Qwen3.8-Flash-Next leads gpt-oss-120b on 2 of the 2 benchmarks they both report (Humanity's Last Exam, GPQA Diamond). gpt-oss-120b shipped 386 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

gpt-oss-120b has 117B parameters, while Qwen3.8-Flash-Next has 125B. Context windows are 128k (gpt-oss-120b) vs 262k (Qwen3.8-Flash-Next).

On Humanity's Last Exam · no tools, Qwen3.8-Flash-Next leads at 35.9% vs gpt-oss-120b at 14.9%. On GPQA Diamond, Qwen3.8-Flash-Next leads at 91.7% vs gpt-oss-120b at 80.1%.

gpt-oss-120b was released by OpenAI on Aug 5 2025.

Qwen3.8-Flash-Next was released by Qwen on Aug 26 2026.

Qwen3.8-Flash-Next leads on Humanity's Last Exam · no tools — gpt-oss-120b 14.9% vs Qwen3.8-Flash-Next 35.9%.

gpt-oss-120b has a 128k context window; Qwen3.8-Flash-Next has 262k.