AI Model Release Tracker - Analytics

Claude Fable 5vsGrok 4.20 Beta

Claude Fable 5
Grok 4.20 Beta
Specifications
Context window
1M
Benchmarks
Nonsense detection
BullshitBench v2
54%
56%Best
Agentic coding
SWE-Bench Pro
80.3%
Coding
SWE-Bench Verified
95.5%
Agentic coding
CursorBench v3.2
70.5%
Agentic coding
CursorBench v3.1
72.9%
Agentic coding
DeepSWE 1.1
70%
Agentic coding
DeepSWE 1.0
66.1%
Next.js coding
Next.js Evals
92%
Agentic computer work
Frontier-Bench v0.1
33.8%
Agentic terminal coding
Terminal-Bench 2.1
88%
Web browsing
BrowseComp
86.9%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
64.5%
Abstract reasoning
ARC-AGI-2
53.3%
Agentic computer use
OSWorld-Verified
85%
Agentic legal work
Harvey's Legal Agent Benchmark
11.25%
Tax questions
TaxEval v2
76.94%
Medical admin work
MedScribe
88.52%
Knowledge work
GDPval-AA
1932
Knowledge work
GDPval-AA v2
1760
Community preference
Arena Elo (Text)
1509Best
1475
Community preference (code)
Arena Elo (Code)
1631
Overview
CompanyAnthropicSpaceXAI
Release dateJun 9 2026Feb 17 2026
AccessProprietaryProprietary

Which is better: Claude Fable 5 or Grok 4.20 Beta?

Claude Fable 5 and Grok 4.20 Beta are evenly matched across the 2 benchmarks they both report (BullshitBench v2, Arena Elo (Text)). Grok 4.20 Beta shipped 112 days before Claude Fable 5, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Grok 4.20 Beta leads at 56% vs Claude Fable 5 at 54%. On Arena Elo (Text), Claude Fable 5 leads at 1509 vs Grok 4.20 Beta at 1475.

Frequently asked questions

Claude Fable 5 was released by Anthropic on Jun 9 2026.

Grok 4.20 Beta was released by SpaceXAI on Feb 17 2026.

Other comparisons