Muse Spark 1.1vsGrok 4.7

Muse Spark 1.1
Grok 4.7
Specifications
Context window
500k
API pricing
Input price
$1.25$2.00
Output price
$4.25$6.00
Cached input price
$0.50
Benchmarks
DeepSWE 1.1
53.3%71%
Harvey's Legal Agent Benchmark
20%19.6%
Benchmarks
SWE-Bench Pro
61.5%
Terminal-Bench 4.0
38%
Terminal-Bench 2.1
80%
MCP Atlas
88.1%
JobBench
54.7%
Toolathlon-Verified
75.6%
Humanity's Last Exam · with tools
62.1%
EEBench
64%
OSWorld-Verified
80.8%
Finance Agent v2
57.2%
TaxEval v2
79.72%
HealthBench Professional
56.7%
MedScribe
88.89%
GDPval-AA v2.1
1695
AA-Briefcase v1.1
1657
CharXiv Reasoning
88.4%
BabyVision
76.3%
Overview
CompanyMetaSpaceXAI
Release dateJul 9 2026Sep 21 2026
AccessProprietaryProprietary

Other comparisons

Muse Spark 1.1vsClaude Fable 5.1Grok 4.7vsClaude Fable 5.1Muse Spark 1.1vsGPT-6 AstraGrok 4.7vsGPT-6 AstraMuse Spark 1.1vsGemini 3.8 FlashGrok 4.7vsGemini 3.8 FlashMuse Spark 1.1vsDeepSeek-V4.1-FlashGrok 4.7vsDeepSeek-V4.1-FlashMuse Spark 1.1vsMistral Medium 3.5Grok 4.7vsMistral Medium 3.5Muse Spark 1.1vsKimi K3Grok 4.7vsKimi K3

Frequently asked questions

Muse Spark 1.1 and Grok 4.7 are evenly matched across the 2 benchmarks they both report (DeepSWE 1.1, Harvey's Legal Agent Benchmark). Muse Spark 1.1 is cheaper on both input and output: $1.25 vs $2.00 per million input tokens, and $4.25 vs $6.00 per million output tokens. Figures are base-tier rates. Muse Spark 1.1 shipped 74 days before Grok 4.7, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On DeepSWE 1.1, Grok 4.7 leads at 71% vs Muse Spark 1.1 at 53.3%. On Harvey's Legal Agent Benchmark, Muse Spark 1.1 leads at 20% vs Grok 4.7 at 19.6%.