Claude 3.5 SonnetvsGrok 4.6

Claude 3.5 Sonnet
Grok 4.6
Benchmarks
Nonsense detection
BullshitBench v2
45%
Coding
SWE-Bench Verified
33.4%
Agentic coding
CursorBench v3.2
69.9%
Agentic coding
DeepSWE 1.1
65.9%
Agentic coding
FrontierCode v1.1 (Extended) · extended split
61.3%
Expert software engineering
APEX-SWE
56.4%
Agentic terminal coding
Terminal-Bench 3.0
26%
Expert agentic work
APEX-Agents
57.5%
Science
GPQA Diamond
59.4%
Agentic legal work
Harvey's Legal Agent Benchmark
15.8%
Overall intelligence
AA Intelligence Index
61
Knowledge work
GDPval-AA v2
1753
Knowledge work
AA-Briefcase
1577
Overview
CompanyAnthropicSpaceXAI
Release dateJun 20 2024Aug 12 2026
AccessProprietaryProprietary

Which is better: Claude 3.5 Sonnet or Grok 4.6?

Claude 3.5 Sonnet and Grok 4.6 don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Claude 3.5 Sonnet shipped 783 days before Grok 4.6, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

Direct benchmark comparisons are unavailable — Claude 3.5 Sonnet and Grok 4.6 don't publish scores on any of the same benchmarks.

Frequently asked questions

Claude 3.5 Sonnet was released by Anthropic on Jun 20 2024.

Grok 4.6 was released by SpaceXAI on Aug 12 2026.

Other comparisons