Claude Sonnet 4.6vsKimi K3

Claude Sonnet 4.6
Kimi K3
Specifications
Parameters
2.8T
Context window
1M
API pricing
Input price
$3.00$3.00
Output price
$15.00$15.00
Cached input price
$0.30$0.30
Cheapest input
$3.00Amazon Bedrock$2.125Morph
Cheapest output
$15.00Amazon Bedrock$11.5502Sail Research
Benchmarks
BullshitBench v2
91%74%
DeepSWE 1.1
30%69%
Next.js Evals
45%84%
MCP Atlas
69.5%84.2%
Humanity's Last Exam · no tools
33.2%43.5%
Humanity's Last Exam · with tools
49%56%
GPQA Diamond
89.9%93.5%
CharXiv Reasoning
72.4%84.8%
MMMU-Pro
74.5%81.6%
Benchmarks
ProgramBench
0%
SWE-Bench Verified
79.6%
DeepSWE 1.0
67.5%
Supabase Evals · with skills
78.3%
Supabase Evals · no skills
81.2%
Terminal-Bench 2.1
88.3%
JobBench
52.9%
Toolathlon-Verified
73.2%
BU Bench
62%
BrowseComp
91.2%
ARC-AGI-2
58.3%
OSWorld-Verified
72.5%
Finance Agent v2
51%
GDPval-AA
1676
GDPval-AA v2
1668
Blueprint-Bench 2
6.7%
MRCR v2 (8-needle) · 128k average
84.9%
threejseval
1549
Overview
CompanyAnthropicMoonshot AI
Release dateFeb 17 2026Jul 16 2026
AccessProprietaryOpen Weight

Other comparisons

Claude Sonnet 4.6vsGPT-6 AstraKimi K3vsGPT-6 AstraClaude Sonnet 4.6vsGemini 3.8 FlashKimi K3vsGemini 3.8 FlashClaude Sonnet 4.6vsMuse Spark 1.3Kimi K3vsMuse Spark 1.3Claude Sonnet 4.6vsGrok 4.6Kimi K3vsGrok 4.6Claude Sonnet 4.6vsDeepSeek-V4.1-FlashKimi K3vsDeepSeek-V4.1-FlashClaude Sonnet 4.6vsMistral Medium 3.5Kimi K3vsMistral Medium 3.5

Frequently asked questions

Kimi K3 leads Claude Sonnet 4.6 on 8 of the 9 benchmarks they both report. Both charge $3.00 per million input tokens. Both charge $15.00 per million output tokens. Claude Sonnet 4.6 shipped 149 days before Kimi K3, so benchmark comparisons should account for the intervening progress.

Claude Sonnet 4.6 is proprietary, while Kimi K3 is open weight.

On BullshitBench v2, Claude Sonnet 4.6 leads at 91% vs Kimi K3 at 74%. On DeepSWE 1.1, Kimi K3 leads at 69% vs Claude Sonnet 4.6 at 30%. On Next.js Evals, Kimi K3 leads at 84% vs Claude Sonnet 4.6 at 45%. On MCP Atlas, Kimi K3 leads at 84.2% vs Claude Sonnet 4.6 at 69.5%. On Humanity's Last Exam · no tools, Kimi K3 leads at 43.5% vs Claude Sonnet 4.6 at 33.2%. On Humanity's Last Exam · with tools, Kimi K3 leads at 56% vs Claude Sonnet 4.6 at 49%. On GPQA Diamond, Kimi K3 leads at 93.5% vs Claude Sonnet 4.6 at 89.9%. On CharXiv Reasoning, Kimi K3 leads at 84.8% vs Claude Sonnet 4.6 at 72.4%. On MMMU-Pro, Kimi K3 leads at 81.6% vs Claude Sonnet 4.6 at 74.5%.