Claude Opus 5.5vsKimi K2.5

Claude Opus 5.5
Kimi K2.5
Specifications
Parameters
1T
Context window
1M256k
API pricing
Input price
$4.00$0.60
Output price
$20.00$3.00
Cached input price
$0.20$0.10
Cheapest input
$0.45SiliconFlow
Cheapest output
$2.25SiliconFlow
Benchmarks
Humanity's Last Exam · with tools
67.7%30.1%
Benchmarks
BullshitBench v2
52%
SWE-Bench Verified
76.8%
FrontierCode v1.1 (Main) · main split
54.4%
Next.js Evals
16%
LiveCodeBench
85%
Terminal-Bench 4.0
66.4%
Terminal-Bench-Science 0.1
58.7%
BrowseComp
60.6%
GPQA Diamond
87.6%
OSWorld 2.0
81.8%
AutomationBench
40%
GDPval-AA v2.1
1846
Chartography · with tools
89%
Overview
CompanyAnthropicMoonshot AI
Release dateSep 22 2026Jan 27 2026
AccessProprietaryOpen Weight

Other comparisons

Claude Opus 5.5vsGPT-6 AstraKimi K2.5vsGPT-6 AstraClaude Opus 5.5vsGemini 3.8 FlashKimi K2.5vsGemini 3.8 FlashClaude Opus 5.5vsMuse Spark 1.3Kimi K2.5vsMuse Spark 1.3Claude Opus 5.5vsGrok 4.7Kimi K2.5vsGrok 4.7Claude Opus 5.5vsDeepSeek-V4.1-FlashKimi K2.5vsDeepSeek-V4.1-FlashClaude Opus 5.5vsMistral Medium 3.5Kimi K2.5vsMistral Medium 3.5

Frequently asked questions

Claude Opus 5.5 leads Kimi K2.5 on 1 of the 1 benchmark they both report (Humanity's Last Exam). Kimi K2.5 is cheaper on both input and output: $0.60 vs $4.00 per million input tokens, and $3.00 vs $20.00 per million output tokens. Kimi K2.5 shipped 238 days before Claude Opus 5.5, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Opus 5.5) vs 256k (Kimi K2.5). Claude Opus 5.5 is proprietary, while Kimi K2.5 is open weight.

On Humanity's Last Exam · with tools, Claude Opus 5.5 leads at 67.7% vs Kimi K2.5 at 30.1%.