GPT-5.4vsGLM-5

GPT-5.4
GLM-5
Specifications
Parameters
744B
Benchmarks
Nonsense detection
BullshitBench v2
48%Best
28%
Coding
SWE-Bench Verified
77.8%
Multilingual coding
SWE-Bench Multilingual
73.3%
Next.js coding
Next.js Evals
83%
Agentic terminal coding
Terminal-Bench 2.0
75.1%Best
56.2%
Software engineering
Expert-SWE (Internal)
68.5%
General tool use
Toolathlon
54.6%
Web browsing
BrowseComp
82.7%Best
75.9%
Cybersecurity
CyberGym
79%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
50.4%
Advanced math
FrontierMath · Tier 1–3
47.6%
Advanced math
FrontierMath · Tier 4
27.1%
Science
GPQA Diamond
92.8%Best
86%
Agentic computer use
OSWorld-Verified
75%
Knowledge work
GDPval (win/tie rate)
83%
Community preference
Arena Elo (Text)
1476
Community preference (code)
Arena Elo (Code)
1462Best
1435
Overview
CompanyOpenAIZ.ai
Release dateMar 5 2026Feb 12 2026
AccessProprietaryOpen Weight

Which is better: GPT-5.4 or GLM-5?

GPT-5.4 leads GLM-5 on 5 of the 5 benchmarks they both report. GLM-5 shipped 21 days before GPT-5.4, so benchmark comparisons should account for the intervening progress.

GPT-5.4 is proprietary, while GLM-5 is open weight.

On BullshitBench v2, GPT-5.4 leads at 48% vs GLM-5 at 28%. On Terminal-Bench 2.0, GPT-5.4 leads at 75.1% vs GLM-5 at 56.2%. On BrowseComp, GPT-5.4 leads at 82.7% vs GLM-5 at 75.9%. On GPQA Diamond, GPT-5.4 leads at 92.8% vs GLM-5 at 86%. On Arena Elo (Code), GPT-5.4 leads at 1462 vs GLM-5 at 1435.

Frequently asked questions

GPT-5.4 was released by OpenAI on Mar 5 2026.

GLM-5 was released by Z.ai on Feb 12 2026.

GPT-5.4 leads on Terminal-Bench 2.0 — GPT-5.4 75.1% vs GLM-5 56.2%.

GPT-5.4 leads on GPQA Diamond — GPT-5.4 92.8% vs GLM-5 86%.

GPT-5.4 is a proprietary model released by OpenAI. GLM-5 is an open weight model released by Z.ai.

Other comparisons