# Grok 4.7

Grok 4.7 is an AI model released by SpaceXAI on Sep 21 2026. It has a 500k context window. Tracked results include 64% on EEBench, 1695 on GDPval-AA v2.1 and 1657 on AA-Briefcase v1.1.

## Facts

| Field | Value |
| --- | --- |
| Model | Grok 4.7 |
| Developer | SpaceXAI |
| Release date | Monday, Sep 21 2026 |
| Licensing | Proprietary |
| Context window | 500k |

## API pricing

All rates in USD per 1,000,000 tokens, pay-as-you-go.

| Tier | Input | Cached input | Output |
| --- | --- | --- | --- |
| Under 200K input tokens | $2.00 | $0.50 | $6.00 |
| At least 200K input tokens | $4.00 | $1.00 | $12.00 |

Verified September 21, 2026 against the first-party source: https://docs.x.ai/developers/pricing

## Tracked benchmark scores

| Benchmark | Score | Source | What it measures |
| --- | --- | --- | --- |
| DeepSWE 1.1 | 71% | Lab | Artificial Analysis' independent test of deep, agentic software-engineering work — the AI has to plan and carry out substantial coding tasks end to end. (Version 1.1 of the test.) Higher is better. |
| Terminal-Bench 4.0 | 38% | Lab | Can the AI work in a command-line terminal — running commands and finishing technical setup tasks the way a developer would? Version 4.0 recalibrated how much time, CPU and memory each task gets, removed eight tasks and fixed nineteen, so fewer runs fail for reasons that have nothing to do with the model. Scores are not comparable with earlier versions. Higher is better. |
| EEBench | 64% | Lab | Electrical-engineering problems — the circuit and systems work an engineer would be handed. Reported in xAI's Grok 4.7 launch comparison, which does not say who publishes the test. Higher is better. |
| Harvey's Legal Agent Benchmark | 19.6% | Lab | Harvey's test of whether an AI agent can complete real legal work — drafting and reviewing documents, working with spreadsheets and presentations, and navigating files the way a lawyer's assistant would. Higher is better. |
| HealthBench Professional | 56.7% | Lab | Realistic health conversations graded against detailed rubrics written by physicians — can the AI respond the way a careful medical professional would? Higher is better. |
| GDPval-AA v2.1 | 1695 | Lab | economically valuable knowledge work (v2.1, Crowd-BT Elo fit) |
| AA-Briefcase v1.1 | 1657 | Lab | Artificial Analysis agentic office-work eval (Elo, v1.1 rating fit) |

## About Grok 4.7

Grok 4.7, released September 21, 2026, was the model Elon Musk had trailed for most of the month: a larger successor to Grok 4.6 that he said traded some serving speed for capability, and that he claimed had been trained in part on SpaceX's own engineering data. It shipped with a 500,000-token context window and at Grok 4.6's prices, unchanged in every column — $2 per million input tokens and $6 per million output under 200K, doubling to $4 and $12 once a prompt crosses that threshold, with cached input at $0.50 and $1.00. The launch came in two pieces that did not compare the same models: a post on X setting Grok 4.7 at xHigh against Grok 4.6, GPT-5.6 Sol and Claude Fable 5.1, and a launch page whose charts swapped in GPT-6 Astra, the OpenAI flagship that had shipped three weeks earlier.

The engineering claim was the one xAI built the launch around, and the two surfaces answered it differently. Grok 4.7 scored 64.0% on EEBench, an electrical-engineering set, against 53.0% for Grok 4.6 — eleven points, its largest gain over its predecessor anywhere — and ahead of Claude Fable 5.1's 56.4%, but the launch page's own chart put GPT-6 Astra above it at 69.3%. The knowledge-work rows fell the same way, second to a rival in each case: a GDPval-AA v2.1 Elo of 1695 against 1605 for Grok 4.6 and 1735 for Claude Fable 5.1, and 1657 on AA-Briefcase v1.1 against 1546 and 1678. Where xAI led outright was legal work, at 19.6% on Harvey's Legal Agent Benchmark — up from 15.8% for Grok 4.6 and far above the 2.5% and 6.7% of the models beside it on the X table. Coding and terminal work stayed the softer side: 71.0% on DeepSWE v1.1, footnoted as run at High rather than xHigh effort, sat just behind GPT-5.6 Sol's 72.7%, and 38.0% on Terminal-Bench 4.0 trailed Claude Fable 5.1's 57.9% by twenty points, with 56.7% on HealthBench Professional third of the four columns. As at the Grok 4.6 launch, the figures for rival models were xAI's own presentation rather than each lab's published results.

## Questions and answers

### When was Grok 4.7 released?

Grok 4.7 was released by SpaceXAI on Monday, Sep 21 2026.

### Who made Grok 4.7?

Grok 4.7 was built by SpaceXAI. Elon Musk's AI company building the Grok series of models. Founded in 2023 as xAI, now part of SpaceX. Cursor (Anysphere) and its Composer coding models were acquired in 2026 and are tracked here.

### How much does Grok 4.7 cost?

Grok 4.7 costs $2.00 per million input tokens and $6.00 per million output tokens through the SpaceXAI API. Cached input is $0.50 per million tokens. Those are the rates for the “Under 200K input tokens” tier; 1 other pricing tier is published for this model. Rates are pay-as-you-go API prices verified against SpaceXAI's published pricing on September 21, 2026.

### What benchmark scores did Grok 4.7 get?

Grok 4.7 reports 7 tracked benchmark scores — DeepSWE 1.1: 71%; Terminal-Bench 4.0: 38%; EEBench: 64%; Harvey's Legal Agent Benchmark: 19.6%; HealthBench Professional: 56.7%; GDPval-AA v2.1: 1695; AA-Briefcase v1.1: 1657. Tracked scores may come from lab reports or independent benchmarks; source details accompany the benchmark data. It holds the best score among all models tracked here on EEBench, GDPval-AA v2.1 and AA-Briefcase v1.1.

### What is the context window of Grok 4.7?

Grok 4.7 has a context window of 500k. That is the maximum amount of input plus output the model can hold in a single request.

### Is Grok 4.7 open source?

No. Grok 4.7 is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.

### What came before and after Grok 4.7?

SpaceXAI's previous tracked release was Grok 4.6 on Aug 12 2026, 40 days earlier. It is the most recent SpaceXAI model tracked on AI Release Tracker.


---

Canonical page: https://aireleasetracker.com/model/xai/grok-4.7
Full dataset: https://aireleasetracker.com/llms-full.txt · JSON: https://aireleasetracker.com/models.json
Source: AI Release Tracker (https://aireleasetracker.com). Most benchmark scores come from lab launch material; gathered results identify the leaderboard that published them.
