Grok 4.5vsGLM-5.3

Grok 4.5
GLM-5.3
Specifications
Parameters
743B
API pricing
Input price
$2.00
Output price
$6.00
Cached input price
$0.30
Cheapest input
$0.8727Inceptron
Cheapest output
$3.36Inceptron
Benchmarks
BullshitBench v2
55%72%
DeepSWE 1.1
54%66.9%
Terminal-Bench 4.0
12.42%41.82%
Terminal-Bench 3.0
15.7%28.3%
Terminal-Bench 2.1
83.3%88.2%
GDPval-AA v2
15261769
Benchmarks
Gray Swan IPI · k = 1
13.4%
Gray Swan IPI · k = 10
54.2%
Gray Swan IPI · k = 15
60.8%
SWE-Bench Pro
64.7%
SWE-Bench Multilingual
78%
DeepSWE 1.0
62%
FrontierCode v1.1 (Extended) · extended split
56.6%
APEX-SWE
53.6%
Next.js Evals
65%
Frontier-Bench v0.1
17.8%
APEX-Agents
47.1%
CyberGym
84.5%
ExploitBench
54.4%
ExploitGym · 6-hour budget
130
ExploitGym · 2-hour budget
105
Humanity's Last Exam · with tools
62.5%
ARC-AGI-2
52.64%
Agent's Last Exam · pass@1
28.5%
AutomationBench
48.2%
Harvey's Legal Agent Benchmark
12.92%
MedScribe
86.88%
AA Intelligence Index
56
AA-Briefcase
1313
Overview
CompanySpaceXAIZ.ai
Release dateJul 8 2026Aug 14 2026
AccessProprietaryOpen Weight

Other comparisons

Grok 4.5vsClaude Fable 5.1GLM-5.3vsClaude Fable 5.1Grok 4.5vsGPT-6 AstraGLM-5.3vsGPT-6 AstraGrok 4.5vsGemini 3.8 FlashGLM-5.3vsGemini 3.8 FlashGrok 4.5vsMuse Spark 1.3GLM-5.3vsMuse Spark 1.3Grok 4.5vsDeepSeek-V4.1-FlashGLM-5.3vsDeepSeek-V4.1-FlashGrok 4.5vsMistral Medium 3.5GLM-5.3vsMistral Medium 3.5

Frequently asked questions

GLM-5.3 leads Grok 4.5 on 6 of the 6 benchmarks they both report. Only Grok 4.5 has a verified first-party API price: $2.00 per million input tokens and $6.00 per million output tokens. No pay-as-you-go API rate is tracked for GLM-5.3. Grok 4.5 shipped 37 days before GLM-5.3, so benchmark comparisons should account for the intervening progress.

Grok 4.5 is proprietary, while GLM-5.3 is open weight.

On BullshitBench v2, GLM-5.3 leads at 72% vs Grok 4.5 at 55%. On DeepSWE 1.1, GLM-5.3 leads at 66.9% vs Grok 4.5 at 54%. On Terminal-Bench 4.0, GLM-5.3 leads at 41.82% vs Grok 4.5 at 12.42%. On Terminal-Bench 3.0, GLM-5.3 leads at 28.3% vs Grok 4.5 at 15.7%. On Terminal-Bench 2.1, GLM-5.3 leads at 88.2% vs Grok 4.5 at 83.3%. On GDPval-AA v2, GLM-5.3 leads at 1769 vs Grok 4.5 at 1526.