gpt-oss-120bvsQwen2.5

gpt-oss-120b
Qwen2.5
Specifications
Parameters
117B
72B
Context window
128k
128k
Benchmarks
Nonsense detection
BullshitBench v2
11%
Coding
SWE-Bench Verified
62.4%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
14.9%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
19%
Science
GPQA Diamond
80.1%
General knowledge
MMLU
90%
Overview
CompanyOpenAIQwen
Release dateAug 5 2025Sep 19 2024
AccessOpen WeightOpen Weight

Which is better: gpt-oss-120b or Qwen2.5?

gpt-oss-120b and Qwen2.5 don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Qwen2.5 shipped 320 days before gpt-oss-120b, so benchmark comparisons should account for the intervening progress.

gpt-oss-120b has 117B parameters, while Qwen2.5 has 72B. Context windows are 128k (gpt-oss-120b) vs 128k (Qwen2.5).

Direct benchmark comparisons are unavailable — gpt-oss-120b and Qwen2.5 don't publish scores on any of the same benchmarks.

Frequently asked questions

gpt-oss-120b was released by OpenAI on Aug 5 2025.

Qwen2.5 was released by Qwen on Sep 19 2024.

gpt-oss-120b has a 128k context window; Qwen2.5 has 128k.

Other comparisons