Kimi K2 ThinkingvsQwen3.6

Kimi K2 Thinking
Qwen3.6
Specifications
Parameters
1T
35B
Context window
256k
256k
Benchmarks
Agentic coding
SWE-Bench Pro
49.5%
Coding
SWE-Bench Verified
71.3%
73.4%Best
Multilingual coding
SWE-Bench Multilingual
67.2%
Agentic terminal coding
Terminal-Bench 2.0
51.5%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
21.4%
Science
GPQA Diamond
86%
Chart reasoning
CharXiv Reasoning
78%
Multimodal reasoning
MMMU-Pro
75.3%
Multimodal
MMMU
81.7%
Overview
CompanyMoonshot AIQwen
Release dateNov 6 2025Apr 16 2026
AccessOpen WeightOpen Weight

Which is better: Kimi K2 Thinking or Qwen3.6?

Qwen3.6 leads Kimi K2 Thinking on 1 of the 1 benchmark they both report (SWE-Bench Verified). Kimi K2 Thinking shipped 161 days before Qwen3.6, so benchmark comparisons should account for the intervening progress.

Kimi K2 Thinking has 1T parameters, while Qwen3.6 has 35B. Context windows are 256k (Kimi K2 Thinking) vs 256k (Qwen3.6).

On SWE-Bench Verified, Qwen3.6 leads at 73.4% vs Kimi K2 Thinking at 71.3%.

Frequently asked questions

Kimi K2 Thinking was released by Moonshot AI on Nov 6 2025.

Qwen3.6 was released by Qwen on Apr 16 2026.

Qwen3.6 leads on SWE-Bench Verified — Kimi K2 Thinking 71.3% vs Qwen3.6 73.4%.

Kimi K2 Thinking has a 256k context window; Qwen3.6 has 256k.

Other comparisons