Claude 3.7 SonnetvsGLM-5

Claude 3.7 Sonnet
GLM-5
Specifications
Parameters
744B
Benchmarks
Nonsense detection
BullshitBench v2
49%Best
28%
Coding
SWE-Bench Verified
62.3%
77.8%Best
Multilingual coding
SWE-Bench Multilingual
73.3%
Agentic terminal coding
Terminal-Bench 2.0
56.2%
Web browsing
BrowseComp
75.9%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
50.4%
Science
GPQA Diamond
68%
86%Best
Community preference (code)
Arena Elo (Code)
1435
Overview
CompanyAnthropicZ.ai
Release dateFeb 24 2025Feb 12 2026
AccessProprietaryOpen Weight

Which is better: Claude 3.7 Sonnet or GLM-5?

GLM-5 leads Claude 3.7 Sonnet on 2 of the 3 benchmarks they both report (BullshitBench v2, SWE-Bench Verified, GPQA Diamond). Claude 3.7 Sonnet shipped 353 days before GLM-5, so benchmark comparisons should account for the intervening progress.

Claude 3.7 Sonnet is proprietary, while GLM-5 is open weight.

On BullshitBench v2, Claude 3.7 Sonnet leads at 49% vs GLM-5 at 28%. On SWE-Bench Verified, GLM-5 leads at 77.8% vs Claude 3.7 Sonnet at 62.3%. On GPQA Diamond, GLM-5 leads at 86% vs Claude 3.7 Sonnet at 68%.

Frequently asked questions

Claude 3.7 Sonnet was released by Anthropic on Feb 24 2025.

GLM-5 was released by Z.ai on Feb 12 2026.

GLM-5 leads on SWE-Bench Verified — Claude 3.7 Sonnet 62.3% vs GLM-5 77.8%.

GLM-5 leads on GPQA Diamond — Claude 3.7 Sonnet 68% vs GLM-5 86%.

Claude 3.7 Sonnet is a proprietary model released by Anthropic. GLM-5 is an open weight model released by Z.ai.

Other comparisons