Claude Opus 5vsMuse Spark

Claude Opus 5
Muse Spark
Specifications
Context window
1M
API pricing
Input price
$5.00
Output price
$25.00
Cached input price
$0.50
Cheapest input
$5.00Amazon Bedrock
Cheapest output
$25.00Amazon Bedrock
Benchmarks
Gray Swan IPI · k = 1
0.2%2.9%
Gray Swan IPI · k = 10
1.6%14.3%
Gray Swan IPI · k = 15
2%16.5%
SWE-Bench Pro
79.2%55%
SWE-Bench Verified
96%77.4%
DeepSWE 1.1
68.8%10%
Humanity's Last Exam · with tools
64.7%50.4%
ARC-AGI-2
90.4%42.5%
Benchmarks
BullshitBench v2
73%
SWE-Bench Multilingual
89.5%
SWE-Bench Multimodal
59.4%
FrontierCode v1.1 (Main) · main split
53.4%
Next.js Evals
92%
Supabase Evals · with skills
91.3%
Supabase Evals · no skills
89.9%
Frontier-Bench v0.1
43.3%
Terminal-Bench 4.0
51.82%
Terminal-Bench 2.1
67.3%
Terminal-Bench-Science 0.1
29%
MCP Atlas
82.2%
JobBench
17%
Toolathlon-Verified
49.4%
BrowseComp
90.8%
Humanity's Last Exam · no tools
56.3%
ARC-AGI-3
30.2%
BioMysteryBench · hard
49.4%
BioMysteryBench · human solved
90.1%
GPQA Diamond
89.5%
OSWorld 2.0
70.6%
OSWorld-Verified
53.3%
AutomationBench
26%
Harvey's Legal Agent Benchmark (Held-out)
11.7%
HealthBench Professional
59.8%
GDPval-AA v2
1861
AA-Briefcase
1685
CharXiv Reasoning
88.9%
BabyVision
39.9%
MMMU
80.4%
threejseval
1770
Overview
CompanyAnthropicMeta
Release dateJul 24 2026Apr 8 2026
AccessProprietaryProprietary

Other comparisons

Claude Opus 5vsGPT-6 AstraMuse SparkvsGPT-6 AstraClaude Opus 5vsGemini 3.8 FlashMuse SparkvsGemini 3.8 FlashClaude Opus 5vsGrok 4.6Muse SparkvsGrok 4.6Claude Opus 5vsDeepSeek-V4-Pro-0813Muse SparkvsDeepSeek-V4-Pro-0813Claude Opus 5vsMistral Medium 3.5Muse SparkvsMistral Medium 3.5Claude Opus 5vsKimi K3Muse SparkvsKimi K3

Frequently asked questions

Claude Opus 5 leads Muse Spark on 8 of the 8 benchmarks they both report. Only Claude Opus 5 has a verified first-party API price: $5.00 per million input tokens and $25.00 per million output tokens. No pay-as-you-go API rate is tracked for Muse Spark. Muse Spark shipped 107 days before Claude Opus 5, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On Gray Swan IPI · k = 1, Claude Opus 5 leads at 0.2% vs Muse Spark at 2.9%. On Gray Swan IPI · k = 10, Claude Opus 5 leads at 1.6% vs Muse Spark at 14.3%. On Gray Swan IPI · k = 15, Claude Opus 5 leads at 2% vs Muse Spark at 16.5%. On SWE-Bench Pro, Claude Opus 5 leads at 79.2% vs Muse Spark at 55%. On SWE-Bench Verified, Claude Opus 5 leads at 96% vs Muse Spark at 77.4%. On DeepSWE 1.1, Claude Opus 5 leads at 68.8% vs Muse Spark at 10%. On Humanity's Last Exam · with tools, Claude Opus 5 leads at 64.7% vs Muse Spark at 50.4%. On ARC-AGI-2, Claude Opus 5 leads at 90.4% vs Muse Spark at 42.5%.