Claude Opus 4.7vsGrok 4.7

Claude Opus 4.7
Grok 4.7
Specifications
Context window
1M500k
API pricing
Input price
$5.00$2.00
Output price
$25.00$6.00
Cached input price
$0.50$0.50
Cheapest input
$5.00Amazon Bedrock
Cheapest output
$25.00Amazon Bedrock
Benchmarks
BullshitBench v2
83%
ProgramBench
0%
SWE-Bench Pro
64.3%
SWE-Bench Verified
87.6%
SWE-Bench Multilingual
80.5%
DeepSWE 1.1
71%
Next.js Evals
58%
Terminal-Bench 4.0
38%
Terminal-Bench 2.1
66.1%
Terminal-Bench 2.0
69.4%
MCP Atlas
79.1%
BrowseComp
79.3%
CyberGym
73.1%
Humanity's Last Exam · no tools
46.9%
Humanity's Last Exam · with tools
54.7%
ARC-AGI-2
75.8%
FrontierMath · Tier 1–3
43.8%
FrontierMath · Tier 4
22.9%
EEBench
64%
GPQA Diamond
94.2%
OSWorld-Verified
78%
Finance Agent v2
51.5%
Harvey's Legal Agent Benchmark
19.6%
HealthBench Professional
56.7%
GDPval-AA
1753
GDPval-AA v2.1
1695
AA-Briefcase v1.1
1657
GDPval (win/tie rate)
80.3%
CharXiv Reasoning
82.1%
MMMU-Pro
75.2%
Blueprint-Bench 2
24.5%
MRCR v2 (8-needle) · 128k average
59.3%
Overview
CompanyAnthropicSpaceXAI
Release dateApr 16 2026Sep 21 2026
AccessProprietaryProprietary

Other comparisons

Claude Opus 4.7vsGPT-6 AstraGrok 4.7vsGPT-6 AstraClaude Opus 4.7vsGemini 3.8 FlashGrok 4.7vsGemini 3.8 FlashClaude Opus 4.7vsMuse Spark 1.3Grok 4.7vsMuse Spark 1.3Claude Opus 4.7vsDeepSeek-V4.1-FlashGrok 4.7vsDeepSeek-V4.1-FlashClaude Opus 4.7vsMistral Medium 3.5Grok 4.7vsMistral Medium 3.5Claude Opus 4.7vsKimi K3Grok 4.7vsKimi K3

Frequently asked questions

Claude Opus 4.7 and Grok 4.7 don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Grok 4.7 is cheaper on both input and output: $2.00 vs $5.00 per million input tokens, and $6.00 vs $25.00 per million output tokens. Figures are base-tier rates. Claude Opus 4.7 shipped 158 days before Grok 4.7, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Opus 4.7) vs 500k (Grok 4.7).

Direct benchmark comparisons are unavailable — Claude Opus 4.7 and Grok 4.7 don't publish scores on any of the same benchmarks.