Claude Fable 5vsGrok 4.6

Claude Fable 5
Grok 4.6
Specifications
Context window
1M
API pricing
Input price
$10.00$2.00
Output price
$50.00$6.00
Cached input price
$1.00$0.50
Cheapest input
$10.00Amazon Bedrock$2.20Amazon Bedrock
Cheapest output
$50.00Amazon Bedrock$6.60Amazon Bedrock
Benchmarks
BullshitBench v2
56%66%
DeepSWE 1.1
70%65.9%
Next.js Evals
77%71%
Terminal-Bench 4.0
44.55%20.3%
Harvey's Legal Agent Benchmark
11.25%15.8%
GDPval-AA v2
17601753
threejseval
15521519
Benchmarks
Gray Swan IPI · k = 1
0.4%
Gray Swan IPI · k = 10
2.3%
Gray Swan IPI · k = 15
2.8%
SWE-Bench Pro
80.3%
SWE-Bench Verified
95.5%
SWE-Bench Multilingual
86.6%
SWE-Bench Multimodal
54.1%
DeepSWE 1.0
66.1%
FrontierCode v1.1 (Extended) · extended split
61.3%
APEX-SWE
56.4%
Frontier-Bench v0.1
33.8%
Terminal-Bench 3.0
26%
Terminal-Bench 2.1
88%
Terminal-Bench-Science 0.1
24.7%
APEX-Agents
57.5%
BrowseComp
86.9%
Humanity's Last Exam · with tools
64.5%
OSWorld-Verified
85%
TaxEval v2
76.94%
HealthBench Professional
63.3%
MedScribe
88.52%
AA Intelligence Index
61
GDPval-AA
1932
AA-Briefcase
1577
Overview
CompanyAnthropicSpaceXAI
Release dateJun 9 2026Aug 12 2026
AccessProprietaryProprietary

Other comparisons

Claude Fable 5vsGPT-6 AstraGrok 4.6vsGPT-6 AstraClaude Fable 5vsGemini 3.8 FlashGrok 4.6vsGemini 3.8 FlashClaude Fable 5vsMuse Spark 1.3Grok 4.6vsMuse Spark 1.3Claude Fable 5vsDeepSeek-V4.1-FlashGrok 4.6vsDeepSeek-V4.1-FlashClaude Fable 5vsMistral Medium 3.5Grok 4.6vsMistral Medium 3.5Claude Fable 5vsKimi K3Grok 4.6vsKimi K3

Frequently asked questions

Claude Fable 5 leads Grok 4.6 on 5 of the 7 benchmarks they both report. Grok 4.6 is cheaper on both input and output: $2.00 vs $10.00 per million input tokens, and $6.00 vs $50.00 per million output tokens. Figures are base-tier rates. Claude Fable 5 shipped 64 days before Grok 4.6, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Grok 4.6 leads at 66% vs Claude Fable 5 at 56%. On DeepSWE 1.1, Claude Fable 5 leads at 70% vs Grok 4.6 at 65.9%. On Next.js Evals, Claude Fable 5 leads at 77% vs Grok 4.6 at 71%. On Terminal-Bench 4.0, Claude Fable 5 leads at 44.55% vs Grok 4.6 at 20.3%. On Harvey's Legal Agent Benchmark, Grok 4.6 leads at 15.8% vs Claude Fable 5 at 11.25%. On GDPval-AA v2, Claude Fable 5 leads at 1760 vs Grok 4.6 at 1753. On threejseval, Claude Fable 5 leads at 1552 vs Grok 4.6 at 1519.