Claude Opus 5.5vsKimi K3

Claude Opus 5.5
Kimi K3
Specifications
Parameters
2.8T
Context window
1M1M
API pricing
Input price
$4.00$3.00
Output price
$20.00$15.00
Cached input price
$0.20$0.30
Cheapest input
$1.50InferenceNet
Cheapest output
$7.50InferenceNet
Benchmarks
Humanity's Last Exam · with tools
67.7%56%
Benchmarks
BullshitBench v2
74%
DeepSWE 1.1
69%
DeepSWE 1.0
67.5%
FrontierCode v1.1 (Main) · main split
54.4%
Next.js Evals
84%
Supabase Evals · with skills
87%
Supabase Evals · no skills
84.1%
Terminal-Bench 4.0
66.4%
Terminal-Bench 2.1
88.3%
Terminal-Bench-Science 0.1
58.7%
MCP Atlas
84.2%
JobBench
52.9%
Toolathlon-Verified
73.2%
BrowseComp
91.2%
Humanity's Last Exam · no tools
43.5%
GPQA Diamond
93.5%
OSWorld 2.0
81.8%
AutomationBench
40%
GDPval-AA v2.1
1846
GDPval-AA v2
1668
CharXiv Reasoning
84.8%
Chartography · with tools
89%
MMMU-Pro
81.6%
threejseval
1564
Overview
CompanyAnthropicMoonshot AI
Release dateSep 22 2026Jul 16 2026
AccessProprietaryOpen Weight

Other comparisons

Claude Opus 5.5vsGPT-6 AstraKimi K3vsGPT-6 AstraClaude Opus 5.5vsGemini 3.8 FlashKimi K3vsGemini 3.8 FlashClaude Opus 5.5vsMuse Spark 1.3Kimi K3vsMuse Spark 1.3Claude Opus 5.5vsGrok 4.7Kimi K3vsGrok 4.7Claude Opus 5.5vsDeepSeek-V4.1-FlashKimi K3vsDeepSeek-V4.1-FlashClaude Opus 5.5vsMistral Medium 3.5Kimi K3vsMistral Medium 3.5

Frequently asked questions

Claude Opus 5.5 leads Kimi K3 on 1 of the 1 benchmark they both report (Humanity's Last Exam). Kimi K3 is cheaper on both input and output: $3.00 vs $4.00 per million input tokens, and $15.00 vs $20.00 per million output tokens. Kimi K3 shipped 68 days before Claude Opus 5.5, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Opus 5.5) vs 1M (Kimi K3). Claude Opus 5.5 is proprietary, while Kimi K3 is open weight.

On Humanity's Last Exam · with tools, Claude Opus 5.5 leads at 67.7% vs Kimi K3 at 56%.