Claude Opus 4.7vsKimi K3

Claude Opus 4.7
Kimi K3
Specifications
Parameters
2.8T
Context window
1M1M
API pricing
Input price
$5.00$3.00
Output price
$25.00$15.00
Cached input price
$0.50$0.30
Cheapest input
$5.00Amazon Bedrock$2.55Makora
Cheapest output
$25.00Amazon Bedrock$12.75Makora
Benchmarks
BullshitBench v2
83%73%
Next.js Evals
69%81%
Terminal-Bench 2.1
66.1%88.3%
MCP Atlas
79.1%84.2%
BrowseComp
79.3%91.2%
Humanity's Last Exam · no tools
46.9%43.5%
Humanity's Last Exam · with tools
54.7%56%
GPQA Diamond
94.2%93.5%
CharXiv Reasoning
82.1%84.8%
MMMU-Pro
75.2%81.6%
Benchmarks
SWE-Bench Pro
64.3%
SWE-Bench Verified
87.6%
SWE-Bench Multilingual
80.5%
DeepSWE 1.1
69%
DeepSWE 1.0
67.5%
Supabase Evals · with skills
86.4%
Supabase Evals · no skills
90.9%
Terminal-Bench 2.0
69.4%
JobBench
52.9%
Toolathlon-Verified
73.2%
CyberGym
73.1%
ARC-AGI-2
75.8%
FrontierMath · Tier 1–3
43.8%
FrontierMath · Tier 4
22.9%
OSWorld-Verified
78%
Finance Agent v2
51.5%
GDPval-AA
1753
GDPval-AA v2
1668
GDPval (win/tie rate)
80.3%
Blueprint-Bench 2
24.5%
MRCR v2 (8-needle) · 128k average
59.3%
Overview
CompanyAnthropicMoonshot AI
Release dateApr 16 2026Jul 16 2026
AccessProprietaryOpen Weight

Other comparisons

Claude Opus 4.7vsGPT-5.6 SolKimi K3vsGPT-5.6 SolClaude Opus 4.7vsGemini 3.7 FlashKimi K3vsGemini 3.7 FlashClaude Opus 4.7vsMuse GlimmerKimi K3vsMuse GlimmerClaude Opus 4.7vsGrok 4.6Kimi K3vsGrok 4.6Claude Opus 4.7vsDeepSeek-V4-Pro-0813Kimi K3vsDeepSeek-V4-Pro-0813Claude Opus 4.7vsMistral Medium 3.5Kimi K3vsMistral Medium 3.5

Frequently asked questions

Kimi K3 leads Claude Opus 4.7 on 7 of the 10 benchmarks they both report. Kimi K3 is cheaper on both input and output: $3.00 vs $5.00 per million input tokens, and $15.00 vs $25.00 per million output tokens. Claude Opus 4.7 shipped 91 days before Kimi K3, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Opus 4.7) vs 1M (Kimi K3). Claude Opus 4.7 is proprietary, while Kimi K3 is open weight.

On BullshitBench v2, Claude Opus 4.7 leads at 83% vs Kimi K3 at 73%. On Next.js Evals, Kimi K3 leads at 81% vs Claude Opus 4.7 at 69%. On Terminal-Bench 2.1, Kimi K3 leads at 88.3% vs Claude Opus 4.7 at 66.1%. On MCP Atlas, Kimi K3 leads at 84.2% vs Claude Opus 4.7 at 79.1%. On BrowseComp, Kimi K3 leads at 91.2% vs Claude Opus 4.7 at 79.3%. On Humanity's Last Exam · no tools, Claude Opus 4.7 leads at 46.9% vs Kimi K3 at 43.5%. On Humanity's Last Exam · with tools, Kimi K3 leads at 56% vs Claude Opus 4.7 at 54.7%. On GPQA Diamond, Claude Opus 4.7 leads at 94.2% vs Kimi K3 at 93.5%. On CharXiv Reasoning, Kimi K3 leads at 84.8% vs Claude Opus 4.7 at 82.1%. On MMMU-Pro, Kimi K3 leads at 81.6% vs Claude Opus 4.7 at 75.2%.

Claude Opus 4.7 was released by Anthropic on Apr 16 2026.

Kimi K3 was released by Moonshot AI on Jul 16 2026.

Kimi K3 leads on Next.js Evals — Claude Opus 4.7 69% vs Kimi K3 81%.

Claude Opus 4.7 leads on Humanity's Last Exam · no tools — Claude Opus 4.7 46.9% vs Kimi K3 43.5%.

Kimi K3 is cheaper on both input and output: $3.00 vs $5.00 per million input tokens, and $15.00 vs $25.00 per million output tokens. Rates are pay-as-you-go API prices verified on August 18, 2026.

Claude Opus 4.7 has a 1M context window; Kimi K3 has 1M.

Claude Opus 4.7 is a proprietary model released by Anthropic. Kimi K3 is an open weight model released by Moonshot AI.