Composer 2.5vsGLM-5.2

Composer 2.5
GLM-5.2
Specifications
Parameters
744B
Context window
1M
Benchmarks
Nonsense detection
BullshitBench v2
31%
Agentic coding
SWE-Bench Pro
54%
62.1%Best
Multilingual coding
SWE-Bench Multilingual
79.8%
Agentic coding
CursorBench v3.2
56.1%Best
55%
Agentic coding
CursorBench v3.1
63.2%Best
54.6%
Agentic coding
DeepSWE 1.1
44%
Agentic coding
DeepSWE 1.0
18%
Next.js coding
Next.js Evals
92%Best
88%
Agentic computer work
Frontier-Bench v0.1
5.1%
Agentic terminal coding
Terminal-Bench 2.1
73%
81%Best
Agentic terminal coding
Terminal-Bench 2.0
69.3%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
40.5%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
54.7%
Science
GPQA Diamond
91.2%
Knowledge work
GDPval-AA v2
1514
Community preference (code)
Arena Elo (Code)
1587
Overview
CompanySpaceXAIZ.ai
Release dateMay 18 2026Jun 16 2026
AccessProprietaryOpen Weight

Which is better: Composer 2.5 or GLM-5.2?

Composer 2.5 leads GLM-5.2 on 3 of the 5 benchmarks they both report. Composer 2.5 shipped 29 days before GLM-5.2, so benchmark comparisons should account for the intervening progress.

Composer 2.5 is proprietary, while GLM-5.2 is open weight.

On SWE-Bench Pro, GLM-5.2 leads at 62.1% vs Composer 2.5 at 54%. On CursorBench v3.2, Composer 2.5 leads at 56.1% vs GLM-5.2 at 55%. On CursorBench v3.1, Composer 2.5 leads at 63.2% vs GLM-5.2 at 54.6%. On Next.js Evals, Composer 2.5 leads at 92% vs GLM-5.2 at 88%. On Terminal-Bench 2.1, GLM-5.2 leads at 81% vs Composer 2.5 at 73%.

Frequently asked questions

Composer 2.5 was released by SpaceXAI on May 18 2026.

GLM-5.2 was released by Z.ai on Jun 16 2026.

GLM-5.2 leads on SWE-Bench Pro — Composer 2.5 54% vs GLM-5.2 62.1%.

Composer 2.5 is a proprietary model released by SpaceXAI. GLM-5.2 is an open weight model released by Z.ai.

Other comparisons