gpt-oss-20bvsQwen3.8-Max

gpt-oss-20b
Qwen3.8-Max
Specifications
Parameters
21B
2.4T
Context window
128k
1M
Benchmarks
Agentic coding
SWE-Bench Pro
67.7%
Coding
SWE-Bench Verified
60.7%
Research reproduction
PaperBench
93%
Agentic terminal coding
Terminal-Bench 2.1
86.6%
Professional tool use
JobBench
53.4%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
10.9%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
17.3%
Science
GPQA Diamond
71.5%
General knowledge
MMLU
85.3%
Agentic computer use
OSWorld-Verified
86.1%
Chart reasoning
CharXiv Reasoning
88.4%
Visual reasoning
BabyVision
82%
Community preference
Arena Elo (Text)
1497
Community preference (code)
Arena Elo (Code)
1667
Overview
CompanyOpenAIQwen
Release dateAug 5 2025Aug 3 2026
AccessOpen WeightProprietary

Which is better: gpt-oss-20b or Qwen3.8-Max?

gpt-oss-20b and Qwen3.8-Max don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. gpt-oss-20b shipped 363 days before Qwen3.8-Max, so benchmark comparisons should account for the intervening progress.

gpt-oss-20b has 21B parameters, while Qwen3.8-Max has 2.4T. Context windows are 128k (gpt-oss-20b) vs 1M (Qwen3.8-Max). gpt-oss-20b is open weight, while Qwen3.8-Max is proprietary.

Direct benchmark comparisons are unavailable — gpt-oss-20b and Qwen3.8-Max don't publish scores on any of the same benchmarks.

Frequently asked questions

gpt-oss-20b was released by OpenAI on Aug 5 2025.

Qwen3.8-Max was released by Qwen on Aug 3 2026.

gpt-oss-20b has a 128k context window; Qwen3.8-Max has 1M.

gpt-oss-20b is an open weight model released by OpenAI. Qwen3.8-Max is a proprietary model released by Qwen.

Other comparisons