Qwen3.6vsGLM-5.1

Qwen3.6
GLM-5.1
Specifications
Parameters
35B
744B
Context window
256k
200k
Benchmarks
Nonsense detection
BullshitBench v2
22%
Agentic coding
SWE-Bench Pro
49.5%
Coding
SWE-Bench Verified
73.4%
Multilingual coding
SWE-Bench Multilingual
67.2%
Next.js coding
Next.js Evals
75%
Agentic terminal coding
Terminal-Bench 2.0
51.5%
63.5%Best
Web browsing
BrowseComp
68%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
21.4%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
52.3%
Science
GPQA Diamond
86%
86.2%Best
Chart reasoning
CharXiv Reasoning
78%
Multimodal reasoning
MMMU-Pro
75.3%
Multimodal
MMMU
81.7%
Community preference (code)
Arena Elo (Code)
1518
Overview
CompanyQwenZ.ai
Release dateApr 16 2026Apr 7 2026
AccessOpen WeightOpen Weight

Which is better: Qwen3.6 or GLM-5.1?

GLM-5.1 leads Qwen3.6 on 2 of the 2 benchmarks they both report (Terminal-Bench 2.0, GPQA Diamond). GLM-5.1 shipped 9 days before Qwen3.6, so benchmark comparisons should account for the intervening progress.

Qwen3.6 has 35B parameters, while GLM-5.1 has 744B. Context windows are 256k (Qwen3.6) vs 200k (GLM-5.1).

On Terminal-Bench 2.0, GLM-5.1 leads at 63.5% vs Qwen3.6 at 51.5%. On GPQA Diamond, GLM-5.1 leads at 86.2% vs Qwen3.6 at 86%.

Frequently asked questions

Qwen3.6 was released by Qwen on Apr 16 2026.

GLM-5.1 was released by Z.ai on Apr 7 2026.

GLM-5.1 leads on Terminal-Bench 2.0 — Qwen3.6 51.5% vs GLM-5.1 63.5%.

GLM-5.1 leads on GPQA Diamond — Qwen3.6 86% vs GLM-5.1 86.2%.

Qwen3.6 has a 256k context window; GLM-5.1 has 200k.

Other comparisons