gpt-oss-120bvsQwen3-Coder-Next

gpt-oss-120b
Qwen3-Coder-Next
Specifications
Parameters
117B
80B
Context window
128k
256k
Benchmarks
Nonsense detection
BullshitBench v2
11%
Agentic coding
SWE-Bench Pro
44.3%
Coding
SWE-Bench Verified
62.4%
70.6%Best
Multilingual coding
SWE-Bench Multilingual
62.8%
Agentic terminal coding
Terminal-Bench 2.0
36.2%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
14.9%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
19%
Science
GPQA Diamond
80.1%
General knowledge
MMLU
90%
Overview
CompanyOpenAIQwen
Release dateAug 5 2025Feb 3 2026
AccessOpen WeightOpen Weight

Which is better: gpt-oss-120b or Qwen3-Coder-Next?

Qwen3-Coder-Next leads gpt-oss-120b on 1 of the 1 benchmark they both report (SWE-Bench Verified). gpt-oss-120b shipped 182 days before Qwen3-Coder-Next, so benchmark comparisons should account for the intervening progress.

gpt-oss-120b has 117B parameters, while Qwen3-Coder-Next has 80B. Context windows are 128k (gpt-oss-120b) vs 256k (Qwen3-Coder-Next).

On SWE-Bench Verified, Qwen3-Coder-Next leads at 70.6% vs gpt-oss-120b at 62.4%.

Frequently asked questions

gpt-oss-120b was released by OpenAI on Aug 5 2025.

Qwen3-Coder-Next was released by Qwen on Feb 3 2026.

Qwen3-Coder-Next leads on SWE-Bench Verified — gpt-oss-120b 62.4% vs Qwen3-Coder-Next 70.6%.

gpt-oss-120b has a 128k context window; Qwen3-Coder-Next has 256k.

Other comparisons