Claude Sonnet 5vsKimi K3

Claude Sonnet 5
Kimi K3
Specifications
Parameters
2.8T
Context window
1M
API pricing
Input price
$2.00$3.00
Output price
$10.00$15.00
Cached input price
$0.20$0.30
Cheapest input
$2.00Amazon Bedrock$2.55Makora
Cheapest output
$10.00Amazon Bedrock$12.75Makora
Benchmarks
BullshitBench v2
80%73%
Next.js Evals
77%81%
Supabase Evals · with skills
95.5%86.4%
Supabase Evals · no skills
90.9%90.9%
Terminal-Bench 2.1
80.4%88.3%
BrowseComp
84.7%91.2%
Humanity's Last Exam · no tools
43.2%43.5%
Humanity's Last Exam · with tools
57.4%56%
Benchmarks
Gray Swan IPI · k = 1
0.6%
Gray Swan IPI · k = 10
4.7%
Gray Swan IPI · k = 15
5.9%
SWE-Bench Pro
63.2%
SWE-Bench Verified
85.2%
DeepSWE 1.1
69%
DeepSWE 1.0
67.5%
Frontier-Bench v0.1
14.6%
Terminal-Bench 4.0
12.42%
MCP Atlas
84.2%
JobBench
52.9%
Toolathlon-Verified
73.2%
GPQA Diamond
93.5%
OSWorld-Verified
81.2%
GDPval-AA
1618
GDPval-AA v2
1668
CharXiv Reasoning
84.8%
MMMU-Pro
81.6%
Overview
CompanyAnthropicMoonshot AI
Release dateJun 30 2026Jul 16 2026
AccessProprietaryOpen Weight

Other comparisons

Claude Sonnet 5vsGPT-5.6 SolKimi K3vsGPT-5.6 SolClaude Sonnet 5vsGemini 3.7 FlashKimi K3vsGemini 3.7 FlashClaude Sonnet 5vsMuse GlimmerKimi K3vsMuse GlimmerClaude Sonnet 5vsGrok 4.6Kimi K3vsGrok 4.6Claude Sonnet 5vsDeepSeek-V4-Pro-0813Kimi K3vsDeepSeek-V4-Pro-0813Claude Sonnet 5vsMistral Medium 3.5Kimi K3vsMistral Medium 3.5

Frequently asked questions

Kimi K3 leads Claude Sonnet 5 on 4 of the 8 benchmarks they both report. Claude Sonnet 5 is cheaper on both input and output: $2.00 vs $3.00 per million input tokens, and $10.00 vs $15.00 per million output tokens. Claude Sonnet 5 shipped 16 days before Kimi K3, so benchmark comparisons should account for the intervening progress.

Claude Sonnet 5 is proprietary, while Kimi K3 is open weight.

On BullshitBench v2, Claude Sonnet 5 leads at 80% vs Kimi K3 at 73%. On Next.js Evals, Kimi K3 leads at 81% vs Claude Sonnet 5 at 77%. On Supabase Evals · with skills, Claude Sonnet 5 leads at 95.5% vs Kimi K3 at 86.4%. On Supabase Evals · no skills, both models score 90.9%. On Terminal-Bench 2.1, Kimi K3 leads at 88.3% vs Claude Sonnet 5 at 80.4%. On BrowseComp, Kimi K3 leads at 91.2% vs Claude Sonnet 5 at 84.7%. On Humanity's Last Exam · no tools, Kimi K3 leads at 43.5% vs Claude Sonnet 5 at 43.2%. On Humanity's Last Exam · with tools, Claude Sonnet 5 leads at 57.4% vs Kimi K3 at 56%.

Claude Sonnet 5 was released by Anthropic on Jun 30 2026.

Kimi K3 was released by Moonshot AI on Jul 16 2026.

Kimi K3 leads on Next.js Evals — Claude Sonnet 5 77% vs Kimi K3 81%.

Kimi K3 leads on Humanity's Last Exam · no tools — Claude Sonnet 5 43.2% vs Kimi K3 43.5%.

Claude Sonnet 5 is cheaper on both input and output: $2.00 vs $3.00 per million input tokens, and $10.00 vs $15.00 per million output tokens. Rates are pay-as-you-go API prices verified on August 18, 2026.

Claude Sonnet 5 is a proprietary model released by Anthropic. Kimi K3 is an open weight model released by Moonshot AI.