Kimi K3vsQwen2.5

Kimi K3
Qwen2.5
Specifications
Parameters
2.8T
72B
Context window
1M
128k
Benchmarks
Nonsense detection
BullshitBench v2
73%
Agentic coding
DeepSWE 1.1
69%
Agentic coding
DeepSWE 1.0
67.5%
Next.js coding
Next.js Evals
92%
Supabase coding
Supabase Evals · with skills
94.7%
Supabase coding
Supabase Evals · no skills
100%
Agentic terminal coding
Terminal-Bench 2.1
88.3%
Multi-step tool use
MCP Atlas
84.2%
Professional tool use
JobBench
52.9%
Personal tool use
Toolathlon-Verified
73.2%
Web browsing
BrowseComp
91.2%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
43.5%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
56%
Science
GPQA Diamond
93.5%
Knowledge work
GDPval-AA v2
1668
Chart reasoning
CharXiv Reasoning
84.8%
Multimodal reasoning
MMMU-Pro
81.6%
Community preference
Arena Elo (Text)
1486
Community preference (code)
Arena Elo (Code)
1679
Overview
CompanyMoonshot AIQwen
Release dateJul 16 2026Sep 19 2024
AccessOpen WeightOpen Weight

Which is better: Kimi K3 or Qwen2.5?

Kimi K3 and Qwen2.5 don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Qwen2.5 shipped 665 days before Kimi K3, so benchmark comparisons should account for the intervening progress.

Kimi K3 has 2.8T parameters, while Qwen2.5 has 72B. Context windows are 1M (Kimi K3) vs 128k (Qwen2.5).

Direct benchmark comparisons are unavailable — Kimi K3 and Qwen2.5 don't publish scores on any of the same benchmarks.

Frequently asked questions

Kimi K3 was released by Moonshot AI on Jul 16 2026.

Qwen2.5 was released by Qwen on Sep 19 2024.

Kimi K3 has a 1M context window; Qwen2.5 has 128k.

Other comparisons