Claude Opus 4.7vsGrok 4.6

Claude Opus 4.7
Grok 4.6
Specifications
Context window
1M
Benchmarks
Nonsense detection
BullshitBench v2
83%
Agentic coding
SWE-Bench Pro
64.3%
Coding
SWE-Bench Verified
87.6%
Multilingual coding
SWE-Bench Multilingual
80.5%
Agentic coding
CursorBench v3.2
69.9%
Agentic coding
CursorBench v3.1
64.8%
Agentic coding
DeepSWE 1.1
65.9%
Agentic coding
FrontierCode v1.1 (Extended) · extended split
61.3%
Expert software engineering
APEX-SWE
56.4%
Next.js coding
Next.js Evals
75%
Agentic terminal coding
Terminal-Bench 3.0
26%
Agentic terminal coding
Terminal-Bench 2.1
66.1%
Agentic terminal coding
Terminal-Bench 2.0
69.4%
Expert agentic work
APEX-Agents
57.5%
Multi-step tool use
MCP Atlas
79.1%
Web browsing
BrowseComp
79.3%
Cybersecurity
CyberGym
73.1%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
46.9%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
54.7%
Abstract reasoning
ARC-AGI-2
75.8%
Advanced math
FrontierMath · Tier 1–3
43.8%
Advanced math
FrontierMath · Tier 4
22.9%
Science
GPQA Diamond
94.2%
Agentic computer use
OSWorld-Verified
78%
Agentic financial analysis
Finance Agent v2
51.5%
Agentic legal work
Harvey's Legal Agent Benchmark
15.8%
Overall intelligence
AA Intelligence Index
61
Knowledge work
GDPval-AA
1753
Knowledge work
GDPval-AA v2
1753
Knowledge work
AA-Briefcase
1577
Knowledge work
GDPval (win/tie rate)
80.3%
Chart reasoning
CharXiv Reasoning
82.1%
Multimodal reasoning
MMMU-Pro
75.2%
Spatial reasoning
Blueprint-Bench 2
24.5%
Long context
MRCR v2 (8-needle) · 128k average
59.3%
Community preference
Arena Elo (Text)
1503
Community preference (code)
Arena Elo (Code)
1557
Overview
CompanyAnthropicSpaceXAI
Release dateApr 16 2026Aug 12 2026
AccessProprietaryProprietary

Which is better: Claude Opus 4.7 or Grok 4.6?

Claude Opus 4.7 and Grok 4.6 don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Claude Opus 4.7 shipped 118 days before Grok 4.6, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

Direct benchmark comparisons are unavailable — Claude Opus 4.7 and Grok 4.6 don't publish scores on any of the same benchmarks.

Frequently asked questions

Claude Opus 4.7 was released by Anthropic on Apr 16 2026.

Grok 4.6 was released by SpaceXAI on Aug 12 2026.

Other comparisons