Kimi K3vsGPT-5.2

Kimi K3
GPT-5.2
Specifications
Parameters
2.8T
Context window
1M
API pricing
Input price
$3.00$1.75
Output price
$15.00$14.00
Cached input price
$0.30$0.175
Cheapest input
$2.10InferenceNet$1.75Azure
Cheapest output
$10.95InferenceNet$14.00Azure
Benchmarks
BullshitBench v2
74%38%
BrowseComp
91.2%65.8%
GPQA Diamond
93.5%92.4%
Benchmarks
SWE-Bench Verified
80%
DeepSWE 1.1
69%
DeepSWE 1.0
67.5%
Next.js Evals
84%
Supabase Evals · with skills
84.1%
Supabase Evals · no skills
78.3%
Terminal-Bench 2.1
88.3%
MCP Atlas
84.2%
JobBench
52.9%
Toolathlon-Verified
73.2%
Humanity's Last Exam · no tools
43.5%
Humanity's Last Exam · with tools
56%
ARC-AGI-2
52.9%
OSWorld-Verified
47.3%
GDPval-AA v2
1668
CharXiv Reasoning
84.8%
MMMU-Pro
81.6%
threejseval
1558
Overview
CompanyMoonshot AIOpenAI
Release dateJul 16 2026Dec 11 2025
AccessOpen WeightProprietary

Other comparisons

Kimi K3vsClaude Fable 5.1GPT-5.2vsClaude Fable 5.1Kimi K3vsGemini 3.8 FlashGPT-5.2vsGemini 3.8 FlashKimi K3vsMuse Spark 1.3GPT-5.2vsMuse Spark 1.3Kimi K3vsGrok 4.6GPT-5.2vsGrok 4.6Kimi K3vsDeepSeek-V4.1-FlashGPT-5.2vsDeepSeek-V4.1-FlashKimi K3vsMistral Medium 3.5GPT-5.2vsMistral Medium 3.5

Frequently asked questions

Kimi K3 leads GPT-5.2 on 3 of the 3 benchmarks they both report (BullshitBench v2, BrowseComp, GPQA Diamond). GPT-5.2 is cheaper on both input and output: $1.75 vs $3.00 per million input tokens, and $14.00 vs $15.00 per million output tokens. GPT-5.2 shipped 217 days before Kimi K3, so benchmark comparisons should account for the intervening progress.

Kimi K3 is open weight, while GPT-5.2 is proprietary.

On BullshitBench v2, Kimi K3 leads at 74% vs GPT-5.2 at 38%. On BrowseComp, Kimi K3 leads at 91.2% vs GPT-5.2 at 65.8%. On GPQA Diamond, Kimi K3 leads at 93.5% vs GPT-5.2 at 92.4%.