Gemini 3.0 FlashvsKimi K2 Thinking

Gemini 3.0 Flash
Kimi K2 Thinking
Specifications
Parameters
1T
Context window
256k
API pricing
Cheapest input
$0.60Google
Cheapest output
$2.50Google
Benchmarks
SWE-Bench Verified
78%71.3%
Benchmarks
SWE-Bench Pro
49.6%
Terminal-Bench 2.1
58%
MCP Atlas
62%
Toolathlon
49.4%
Humanity's Last Exam · no tools
33.7%
ARC-AGI-2
33.6%
GPQA Diamond
90.4%
OSWorld-Verified
65.1%
Finance Agent v2
42.6%
GDPval-AA
1204
CharXiv Reasoning
80.3%
MMMU-Pro
81.2%
Blueprint-Bench 2
0%
MRCR v2 (8-needle) · 128k average
67.2%
MRCR v2 (8-needle) · 1M pointwise
22.1%
Overview
CompanyGoogleMoonshot AI
Release dateDec 17 2025Nov 6 2025
AccessProprietaryOpen Weight

Other comparisons

Gemini 3.0 FlashvsClaude Opus 5Kimi K2 ThinkingvsClaude Opus 5Gemini 3.0 FlashvsGPT-5.6 SolKimi K2 ThinkingvsGPT-5.6 SolGemini 3.0 FlashvsMuse GlimmerKimi K2 ThinkingvsMuse GlimmerGemini 3.0 FlashvsGrok 4.6Kimi K2 ThinkingvsGrok 4.6Gemini 3.0 FlashvsDeepSeek-V4-Pro-0813Kimi K2 ThinkingvsDeepSeek-V4-Pro-0813Gemini 3.0 FlashvsMistral Medium 3.5Kimi K2 ThinkingvsMistral Medium 3.5

Frequently asked questions

Gemini 3.0 Flash leads Kimi K2 Thinking on 1 of the 1 benchmark they both report (SWE-Bench Verified). Kimi K2 Thinking shipped 41 days before Gemini 3.0 Flash, so benchmark comparisons should account for the intervening progress.

Gemini 3.0 Flash is proprietary, while Kimi K2 Thinking is open weight.

On SWE-Bench Verified, Gemini 3.0 Flash leads at 78% vs Kimi K2 Thinking at 71.3%.

Gemini 3.0 Flash was released by Google on Dec 17 2025.

Kimi K2 Thinking was released by Moonshot AI on Nov 6 2025.

Gemini 3.0 Flash leads on SWE-Bench Verified — Gemini 3.0 Flash 78% vs Kimi K2 Thinking 71.3%.

Gemini 3.0 Flash is a proprietary model released by Google. Kimi K2 Thinking is an open weight model released by Moonshot AI.