gpt-oss-120bvsQwen3.7-Max

gpt-oss-120b
Qwen3.7-Max
Benchmarks
Nonsense detection
BullshitBench v2
11%
71%Best
Agentic coding
SWE-Bench Pro
60.6%
Agentic terminal coding
Terminal-Bench 2.0
69.7%
Multi-step tool use
MCP Atlas
76.4%
Science
GPQA Diamond
92.4%
Overview
CompanyOpenAIQwen
Release dateAug 5 2025May 20 2026
AccessOpen WeightProprietary

Which is better: gpt-oss-120b or Qwen3.7-Max?

Qwen3.7-Max leads gpt-oss-120b on 1 of the 1 benchmark they both report (BullshitBench v2). gpt-oss-120b shipped 288 days before Qwen3.7-Max, so benchmark comparisons should account for the intervening progress.

gpt-oss-120b is open weight, while Qwen3.7-Max is proprietary.

On BullshitBench v2, Qwen3.7-Max leads at 71% vs gpt-oss-120b at 11%.

Frequently asked questions

gpt-oss-120b was released by OpenAI on Aug 5 2025.

Qwen3.7-Max was released by Qwen on May 20 2026.

gpt-oss-120b is an open weight model released by OpenAI. Qwen3.7-Max is a proprietary model released by Qwen.

Other comparisons