DeepSeek-V4-Pro-0813vsKimi K3

DeepSeek-V4-Pro-0813
Kimi K3
Specifications
Parameters
2.8T
Context window
1M
Benchmarks
Nonsense detection
BullshitBench v2
73%
Agentic coding
DeepSWE 1.1
69%
Agentic coding
DeepSWE 1.0
67.5%
Next.js coding
Next.js Evals
92%
Supabase coding
Supabase Evals · with skills
86.4%
Supabase coding
Supabase Evals · no skills
90.9%
Agentic terminal coding
Terminal-Bench 2.1
87.9%
88.3%Best
Multi-step tool use
MCP Atlas
84.2%
Professional tool use
JobBench
52.9%
Personal tool use
Toolathlon-Verified
74.1%Best
73.2%
Web browsing
BrowseComp
91.2%
Cybersecurity
CyberGym
83.3%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
42.7%
43.5%Best
Multidisciplinary reasoning
Humanity's Last Exam · with tools
60%Best
56%
Science
GPQA Diamond
93.5%
Business workflows
AutomationBench
31.8%
Knowledge work
GDPval-AA v2
1668
Chart reasoning
CharXiv Reasoning
84.8%
Multimodal reasoning
MMMU-Pro
81.6%
Community preference
Arena Elo (Text)
1486
Community preference (code)
Arena Elo (Code)
1679
Overview
CompanyDeepSeekMoonshot AI
Release dateAug 13 2026Jul 16 2026
AccessProprietaryOpen Weight

Which is better: DeepSeek-V4-Pro-0813 or Kimi K3?

DeepSeek-V4-Pro-0813 and Kimi K3 are evenly matched across the 4 benchmarks they both report (Terminal-Bench 2.1, Toolathlon-Verified, Humanity's Last Exam). Kimi K3 shipped 28 days before DeepSeek-V4-Pro-0813, so benchmark comparisons should account for the intervening progress.

DeepSeek-V4-Pro-0813 is proprietary, while Kimi K3 is open weight.

On Terminal-Bench 2.1, Kimi K3 leads at 88.3% vs DeepSeek-V4-Pro-0813 at 87.9%. On Toolathlon-Verified, DeepSeek-V4-Pro-0813 leads at 74.1% vs Kimi K3 at 73.2%. On Humanity's Last Exam · no tools, Kimi K3 leads at 43.5% vs DeepSeek-V4-Pro-0813 at 42.7%. On Humanity's Last Exam · with tools, DeepSeek-V4-Pro-0813 leads at 60% vs Kimi K3 at 56%.

Frequently asked questions

DeepSeek-V4-Pro-0813 was released by DeepSeek on Aug 13 2026.

Kimi K3 was released by Moonshot AI on Jul 16 2026.

Kimi K3 leads on Terminal-Bench 2.1 — DeepSeek-V4-Pro-0813 87.9% vs Kimi K3 88.3%.

Kimi K3 leads on Humanity's Last Exam · no tools — DeepSeek-V4-Pro-0813 42.7% vs Kimi K3 43.5%.

DeepSeek-V4-Pro-0813 is a proprietary model released by DeepSeek. Kimi K3 is an open weight model released by Moonshot AI.

Other comparisons