Claude Sonnet 4.5vsGrok 4.6

Claude Sonnet 4.5
Grok 4.6
Benchmarks
Nonsense detection
BullshitBench v2
79%
Coding
SWE-Bench Verified
77.2%
Agentic coding
CursorBench v3.2
69.9%
Agentic coding
DeepSWE 1.1
65.9%
Agentic coding
FrontierCode v1.1 (Extended) · extended split
61.3%
Expert software engineering
APEX-SWE
56.4%
Next.js coding
Next.js Evals
50%
Agentic terminal coding
Terminal-Bench 3.0
26%
Expert agentic work
APEX-Agents
57.5%
Science
GPQA Diamond
83.4%
Agentic legal work
Harvey's Legal Agent Benchmark
15.8%
Overall intelligence
AA Intelligence Index
61
Knowledge work
GDPval-AA v2
1753
Knowledge work
AA-Briefcase
1577
Multimodal
MMMU
68%
Overview
CompanyAnthropicSpaceXAI
Release dateSep 29 2025Aug 12 2026
AccessProprietaryProprietary

Which is better: Claude Sonnet 4.5 or Grok 4.6?

Claude Sonnet 4.5 and Grok 4.6 don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Claude Sonnet 4.5 shipped 317 days before Grok 4.6, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

Direct benchmark comparisons are unavailable — Claude Sonnet 4.5 and Grok 4.6 don't publish scores on any of the same benchmarks.

Frequently asked questions

Claude Sonnet 4.5 was released by Anthropic on Sep 29 2025.

Grok 4.6 was released by SpaceXAI on Aug 12 2026.

Other comparisons