gpt-oss-120bvsQwen3.6-Plus
gpt-oss-120b | Qwen3.6-Plus | |
|---|---|---|
| Benchmarks | ||
Nonsense detection BullshitBench v2Given a confidently-worded but nonsensical prompt, does the AI spot that it makes no sense and push back — instead of playing along and inventing an answer? The score is how often it clearly called out the nonsense. Higher is better. | 11% | 72%Best |
| Overview | ||
| Company | OpenAI | Qwen |
| Release date | Aug 5 2025 | Apr 2 2026 |
| Access | Open Weight | Proprietary |
Which is better: gpt-oss-120b or Qwen3.6-Plus?
Qwen3.6-Plus leads gpt-oss-120b on 1 of the 1 benchmark they both report (BullshitBench v2). gpt-oss-120b shipped 240 days before Qwen3.6-Plus, so benchmark comparisons should account for the intervening progress.
gpt-oss-120b is open weight, while Qwen3.6-Plus is proprietary.
On BullshitBench v2, Qwen3.6-Plus leads at 72% vs gpt-oss-120b at 11%.
Frequently asked questions
gpt-oss-120b was released by OpenAI on Aug 5 2025.
Qwen3.6-Plus was released by Qwen on Apr 2 2026.
gpt-oss-120b is an open weight model released by OpenAI. Qwen3.6-Plus is a proprietary model released by Qwen.
Other comparisons
gpt-oss-120bvsClaude Opus 5Qwen3.6-PlusvsClaude Opus 5gpt-oss-120bvsGemini 3.6 FlashQwen3.6-PlusvsGemini 3.6 Flashgpt-oss-120bvsMuse Spark 1.2Qwen3.6-PlusvsMuse Spark 1.2gpt-oss-120bvsGrok 4.5Qwen3.6-PlusvsGrok 4.5gpt-oss-120bvsDeepSeek-V4-Flash-0731Qwen3.6-PlusvsDeepSeek-V4-Flash-0731gpt-oss-120bvsMistral Medium 3.5Qwen3.6-PlusvsMistral Medium 3.5