Claude Fable 5vsMuse Spark 1.2

Claude Fable 5
Muse Spark 1.2
Specifications
Context window
1M
1M
Benchmarks
Nonsense detection
BullshitBench v2
54%
Prompt injection robustness
Gray Swan IPI · k = 1
0.4%
Prompt injection robustness
Gray Swan IPI · k = 10
2.3%
Prompt injection robustness
Gray Swan IPI · k = 15
2.8%
Agentic coding
SWE-Bench Pro
80.3%
Coding
SWE-Bench Verified
95.5%
Agentic coding
CursorBench v3.2
70.5%
Agentic coding
CursorBench v3.1
72.9%
Agentic coding
DeepSWE 1.1
70%Best
59.3%
Agentic coding
DeepSWE 1.0
66.1%
Next.js coding
Next.js Evals
92%
Agentic computer work
Frontier-Bench v0.1
33.8%
Agentic terminal coding
Terminal-Bench 2.1
88%Best
82.9%
Web browsing
BrowseComp
86.9%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
64.5%
Agentic computer use
OSWorld-Verified
85%
Agentic legal work
Harvey's Legal Agent Benchmark
11.25%
Tax questions
TaxEval v2
76.94%
Medical admin work
MedScribe
88.52%
Knowledge work
GDPval-AA
1932
Knowledge work
GDPval-AA v2
1760
Community preference
Arena Elo (Text)
1509
Community preference (code)
Arena Elo (Code)
1631
Overview
CompanyAnthropicMeta
Release dateJun 9 2026Aug 5 2026
AccessProprietaryProprietary

Which is better: Claude Fable 5 or Muse Spark 1.2?

Claude Fable 5 leads Muse Spark 1.2 on 2 of the 2 benchmarks they both report (DeepSWE 1.1, Terminal-Bench 2.1). Claude Fable 5 shipped 57 days before Muse Spark 1.2, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Fable 5) vs 1M (Muse Spark 1.2).

On DeepSWE 1.1, Claude Fable 5 leads at 70% vs Muse Spark 1.2 at 59.3%. On Terminal-Bench 2.1, Claude Fable 5 leads at 88% vs Muse Spark 1.2 at 82.9%.

Frequently asked questions

Claude Fable 5 was released by Anthropic on Jun 9 2026.

Muse Spark 1.2 was released by Meta on Aug 5 2026.

Claude Fable 5 leads on DeepSWE 1.1 — Claude Fable 5 70% vs Muse Spark 1.2 59.3%.

Claude Fable 5 has a 1M context window; Muse Spark 1.2 has 1M.

Other comparisons