Claude Sonnet 4.6vsMistral Medium 3.5

Claude Sonnet 4.6
Mistral Medium 3.5
Specifications
Parameters
128B
Context window
256k
API pricing
Input price
$3.00$1.50
Output price
$15.00$7.50
Cached input price
$0.30
Cheapest input
$3.00Amazon Bedrock
Cheapest output
$15.00Amazon Bedrock
Benchmarks
SWE-Bench Verified
79.6%77.6%
Benchmarks
BullshitBench v2
91%
ProgramBench
0%
DeepSWE 1.1
30%
Next.js Evals
45%
MCP Atlas
69.5%
BU Bench
62%
Humanity's Last Exam · no tools
33.2%
Humanity's Last Exam · with tools
49%
ARC-AGI-2
58.3%
GPQA Diamond
89.9%
OSWorld-Verified
72.5%
Finance Agent v2
51%
GDPval-AA
1676
CharXiv Reasoning
72.4%
MMMU-Pro
74.5%
Blueprint-Bench 2
6.7%
MRCR v2 (8-needle) · 128k average
84.9%
Overview
CompanyAnthropicMistral
Release dateFeb 17 2026Apr 29 2026
AccessProprietaryOpen Weight

Other comparisons

Claude Sonnet 4.6vsGPT-6 AstraMistral Medium 3.5vsGPT-6 AstraClaude Sonnet 4.6vsGemini 3.8 FlashMistral Medium 3.5vsGemini 3.8 FlashClaude Sonnet 4.6vsMuse Spark 1.3Mistral Medium 3.5vsMuse Spark 1.3Claude Sonnet 4.6vsGrok 4.6Mistral Medium 3.5vsGrok 4.6Claude Sonnet 4.6vsDeepSeek-V4.1-FlashMistral Medium 3.5vsDeepSeek-V4.1-FlashClaude Sonnet 4.6vsKimi K3Mistral Medium 3.5vsKimi K3

Frequently asked questions

Claude Sonnet 4.6 leads Mistral Medium 3.5 on 1 of the 1 benchmark they both report (SWE-Bench Verified). Mistral Medium 3.5 is cheaper on both input and output: $1.50 vs $3.00 per million input tokens, and $7.50 vs $15.00 per million output tokens. Claude Sonnet 4.6 shipped 71 days before Mistral Medium 3.5, so benchmark comparisons should account for the intervening progress.

Claude Sonnet 4.6 is proprietary, while Mistral Medium 3.5 is open weight.

On SWE-Bench Verified, Claude Sonnet 4.6 leads at 79.6% vs Mistral Medium 3.5 at 77.6%.