Claude 3.7 SonnetvsGLM-4.5

Claude 3.7 Sonnet
GLM-4.5
Specifications
Parameters
355B
Context window
128k
Benchmarks
Nonsense detection
BullshitBench v2
49%Best
8%
Coding
SWE-Bench Verified
62.3%
64.2%Best
Science
GPQA Diamond
68%
79.1%Best
Overview
CompanyAnthropicZ.ai
Release dateFeb 24 2025Jul 28 2025
AccessProprietaryOpen Weight

Which is better: Claude 3.7 Sonnet or GLM-4.5?

GLM-4.5 leads Claude 3.7 Sonnet on 2 of the 3 benchmarks they both report (BullshitBench v2, SWE-Bench Verified, GPQA Diamond). Claude 3.7 Sonnet shipped 154 days before GLM-4.5, so benchmark comparisons should account for the intervening progress.

Claude 3.7 Sonnet is proprietary, while GLM-4.5 is open weight.

On BullshitBench v2, Claude 3.7 Sonnet leads at 49% vs GLM-4.5 at 8%. On SWE-Bench Verified, GLM-4.5 leads at 64.2% vs Claude 3.7 Sonnet at 62.3%. On GPQA Diamond, GLM-4.5 leads at 79.1% vs Claude 3.7 Sonnet at 68%.

Frequently asked questions

Claude 3.7 Sonnet was released by Anthropic on Feb 24 2025.

GLM-4.5 was released by Z.ai on Jul 28 2025.

GLM-4.5 leads on SWE-Bench Verified — Claude 3.7 Sonnet 62.3% vs GLM-4.5 64.2%.

GLM-4.5 leads on GPQA Diamond — Claude 3.7 Sonnet 68% vs GLM-4.5 79.1%.

Claude 3.7 Sonnet is a proprietary model released by Anthropic. GLM-4.5 is an open weight model released by Z.ai.

Other comparisons