GPT-5.4vsGrok 4.7

GPT-5.4
Grok 4.7
Specifications
Context window
500k
API pricing
Input price
$2.50$2.00
Output price
$15.00$6.00
Cached input price
$0.25$0.50
Cheapest input
$2.50Azure
Cheapest output
$15.00Azure
Benchmarks
CyberGym
79%80.3%
Benchmarks
BullshitBench v2
48%
ProgramBench
0%
DeepSWE 1.1
71%
FrontierSWE V2
29%
SWE-Marathon v1.1
46%
Next.js Evals
65%
CADGenBench
44.4%
Terminal-Bench 4.0
38%
Terminal-Bench 2.0
75.1%
Expert-SWE (Internal)
68.5%
Toolathlon
54.6%
BrowseComp
82.7%
CVE-Bench
37.7%
CathedralBench
29%
Humanity's Last Exam · with tools
52.1%
ARC-AGI-2
73.95%
FrontierMath · Tier 1–3
47.6%
FrontierMath · Tier 4
27.1%
LatchBio Capabilities v1.0
44.5%
EEBench
66%
GPQA Diamond
92.8%
OSWorld-Verified
75%
Harvey's Legal Agent Benchmark
19.6%
HealthBench Professional
56.7%
GDPval-AA v2.1
1695
AA-Briefcase v1.1
1657
GDPval (win/tie rate)
83%
Overview
CompanyOpenAISpaceXAI
Release dateMar 5 2026Sep 21 2026
AccessProprietaryProprietary

Other comparisons

GPT-5.4vsClaude Fable 5.1Grok 4.7vsClaude Fable 5.1GPT-5.4vsGemini 3.8 FlashGrok 4.7vsGemini 3.8 FlashGPT-5.4vsMuse Spark 1.3Grok 4.7vsMuse Spark 1.3GPT-5.4vsDeepSeek-V4.1-FlashGrok 4.7vsDeepSeek-V4.1-FlashGPT-5.4vsMistral Medium 3.5Grok 4.7vsMistral Medium 3.5GPT-5.4vsKimi K3Grok 4.7vsKimi K3

Frequently asked questions

Grok 4.7 leads GPT-5.4 on 1 of the 1 benchmark they both report (CyberGym). Grok 4.7 is cheaper on both input and output: $2.00 vs $2.50 per million input tokens, and $6.00 vs $15.00 per million output tokens. Figures are base-tier rates. GPT-5.4 shipped 200 days before Grok 4.7, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On CyberGym, Grok 4.7 leads at 80.3% vs GPT-5.4 at 79%.