GPT-5.6 LunavsQwen3.7-Max

GPT-5.6 Luna
Qwen3.7-Max
Benchmarks
Nonsense detection
BullshitBench v2
40%
71%Best
Prompt injection robustness
Gray Swan IPI · k = 1
8.3%
Prompt injection robustness
Gray Swan IPI · k = 10
38.6%
Prompt injection robustness
Gray Swan IPI · k = 15
43.9%
Agentic coding
SWE-Bench Pro
60.6%
Agentic coding
CursorBench v3.2
61.1%
Agentic computer work
Frontier-Bench v0.1
14.3%
Agentic terminal coding
Terminal-Bench 2.1
82.5%
Agentic terminal coding
Terminal-Bench 2.0
69.7%
Multi-step tool use
MCP Atlas
76.4%
Science
GPQA Diamond
92.4%
Community preference (code)
Arena Elo (Code)
1523
Overview
CompanyOpenAIQwen
Release dateJun 26 2026May 20 2026
AccessProprietaryProprietary

Which is better: GPT-5.6 Luna or Qwen3.7-Max?

Qwen3.7-Max leads GPT-5.6 Luna on 1 of the 1 benchmark they both report (BullshitBench v2). Qwen3.7-Max shipped 37 days before GPT-5.6 Luna, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Qwen3.7-Max leads at 71% vs GPT-5.6 Luna at 40%.

Frequently asked questions

GPT-5.6 Luna was released by OpenAI on Jun 26 2026.

Qwen3.7-Max was released by Qwen on May 20 2026.

Other comparisons