# Grok 4.6

Grok 4.6 is an AI model released by SpaceXAI on Aug 12 2026. At release it scored 61.3% on FrontierCode v1.1 (Extended) (extended split), 56.4% on APEX-SWE and 26% on Terminal-Bench 3.0.

## Facts

| Field | Value |
| --- | --- |
| Model | Grok 4.6 |
| Developer | SpaceXAI |
| Release date | Wednesday, Aug 12 2026 |
| Licensing | Proprietary |

## Benchmark scores published at release

| Benchmark | Score | What it measures |
| --- | --- | --- |
| CursorBench v3.2 | 69.9% | Cursor's own test of harder, real-world coding tasks inside a code editor, on the refreshed v3.2 task set. Scores aren't comparable with v3.1. Higher is better. |
| DeepSWE 1.1 | 65.9% | Artificial Analysis' independent test of deep, agentic software-engineering work — the AI has to plan and carry out substantial coding tasks end to end. (Version 1.1 of the test.) Higher is better. |
| FrontierCode v1.1 (Extended) (extended split) | 61.3% | frontier-difficulty agentic coding tasks (v1.1, extended split) |
| APEX-SWE | 56.4% | expert-level software-engineering tasks (AI Productivity Index) |
| Terminal-Bench 3.0 | 26% | command-line task completion (v3.0, much harder task set) |
| APEX-Agents | 57.5% | expert-level agentic work tasks (AI Productivity Index) |
| Harvey's Legal Agent Benchmark | 15.8% | Harvey's test of whether an AI agent can complete real legal work — drafting and reviewing documents, working with spreadsheets and presentations, and navigating files the way a lawyer's assistant would. Higher is better. |
| AA Intelligence Index | 61 | Artificial Analysis composite intelligence index across evals |
| GDPval-AA v2 | 1753 | economically valuable knowledge work (v2, re-based Elo) |
| AA-Briefcase | 1577 | Artificial Analysis agentic office-work eval (Elo) |

## About Grok 4.6

Grok 4.6, released August 12, 2026, was xAI's bet that post-training alone could deliver a flagship upgrade. It reused the 1.5-trillion-parameter base model of Grok 4.5, five weeks its senior, and poured the gains into supervised fine-tuning on regenerated trajectories and wide-ranging reinforcement learning across engineering and domain-specific environments. At release it scored 61 on the Artificial Analysis Intelligence Index — up from 56 for Grok 4.5, and level with GPT-5.6 Sol at the time — while priced at $2 per million input tokens and $6 per million output.

xAI pitched it at long-running agents and knowledge work rather than raw coding leaderboards. At launch it posted a GDPval-AA v2 Elo of 1753, 15.8% on Harvey's Legal Agent Benchmark — the top score in xAI's own launch comparison — and 69.9% on CursorBench v3.2, within a point of Claude Fable 5's published number at the time. Terminal work stayed a relative weakness, at 26% on the much harder Terminal-Bench 3.0. Elon Musk said the larger Grok 4.7 would follow within weeks, trading some serving speed for capability.

## Questions and answers

### When was Grok 4.6 released?

Grok 4.6 was released by SpaceXAI on Wednesday, Aug 12 2026.

### Who made Grok 4.6?

Grok 4.6 was built by SpaceXAI. Elon Musk's AI company building the Grok series of models. Founded in 2023 as xAI, now part of SpaceX.

### What benchmark scores did Grok 4.6 get?

Grok 4.6 reports 10 tracked benchmark scores — CursorBench v3.2: 69.9%; DeepSWE 1.1: 65.9%; FrontierCode v1.1 (Extended) (extended split): 61.3%; APEX-SWE: 56.4%; Terminal-Bench 3.0: 26%; APEX-Agents: 57.5%; Harvey's Legal Agent Benchmark: 15.8%; AA Intelligence Index: 61; GDPval-AA v2: 1753; AA-Briefcase: 1577. Scores are the figures published at release by SpaceXAI. It holds the best score among all models tracked here on FrontierCode v1.1 (Extended) (extended split), APEX-SWE, Terminal-Bench 3.0, APEX-Agents, AA Intelligence Index and AA-Briefcase.

### Is Grok 4.6 open source?

No. Grok 4.6 is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.

### What came before and after Grok 4.6?

SpaceXAI's previous tracked release was Grok 4.5 on Jul 8 2026, 35 days earlier. It is the most recent SpaceXAI model tracked on AI Release Tracker.


---

Canonical page: https://aireleasetracker.com/model/xai/grok-4.6
Full dataset: https://aireleasetracker.com/llms-full.txt · JSON: https://aireleasetracker.com/models.json
Source: AI Release Tracker (https://aireleasetracker.com). Benchmark scores are the figures published by the releasing lab at launch.
