gpt-oss-120bvsQwen3.8-Max

gpt-oss-120b
Qwen3.8-Max
Specifications
Parameters
117B
2.4T
Context window
128k
1M
Benchmarks
Nonsense detection
BullshitBench v2
11%
Agentic coding
SWE-Bench Pro
67.7%
Coding
SWE-Bench Verified
62.4%
Research reproduction
PaperBench
93%
Agentic terminal coding
Terminal-Bench 2.1
86.6%
Professional tool use
JobBench
53.4%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
14.9%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
19%
Science
GPQA Diamond
80.1%
General knowledge
MMLU
90%
Agentic computer use
OSWorld-Verified
86.1%
Chart reasoning
CharXiv Reasoning
88.4%
Visual reasoning
BabyVision
82%
Community preference
Arena Elo (Text)
1497
Community preference (code)
Arena Elo (Code)
1667
Overview
CompanyOpenAIQwen
Release dateAug 5 2025Aug 3 2026
AccessOpen WeightProprietary

Which is better: gpt-oss-120b or Qwen3.8-Max?

gpt-oss-120b and Qwen3.8-Max don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. gpt-oss-120b shipped 363 days before Qwen3.8-Max, so benchmark comparisons should account for the intervening progress.

gpt-oss-120b has 117B parameters, while Qwen3.8-Max has 2.4T. Context windows are 128k (gpt-oss-120b) vs 1M (Qwen3.8-Max). gpt-oss-120b is open weight, while Qwen3.8-Max is proprietary.

Direct benchmark comparisons are unavailable — gpt-oss-120b and Qwen3.8-Max don't publish scores on any of the same benchmarks.

Frequently asked questions

gpt-oss-120b was released by OpenAI on Aug 5 2025.

Qwen3.8-Max was released by Qwen on Aug 3 2026.

gpt-oss-120b has a 128k context window; Qwen3.8-Max has 1M.

gpt-oss-120b is an open weight model released by OpenAI. Qwen3.8-Max is a proprietary model released by Qwen.

Other comparisons