Kimi K2 ThinkingvsQwen3.8-27B

Kimi K2 Thinking
Qwen3.8-27B
Specifications
Parameters
1T
27B
Context window
256k
262k
Benchmarks
Agentic coding
SWE-Bench Pro
61.7%
Coding
SWE-Bench Verified
71.3%
Agentic coding
DeepSWE 1.1
42.2%
Repo-level code generation
NL2Repo-Bench
42.3%
Software engineering
QwenSWEBench
79%
Competitive coding
LiveCodeBench
90.3%
Agentic terminal coding
Terminal-Bench 2.1
73%
Professional tool use
JobBench
33.4%
Long-horizon office work
CoWorkBench
70.7%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
30.8%
Science
GPQA Diamond
89.2%
Instruction following
IFBench
79.5%
Agentic computer use
Agent's Last Exam · pass@1
20.4%
Agentic computer use
Agent's Last Exam · score
42.9%
Overview
CompanyMoonshot AIQwen
Release dateNov 6 2025Aug 14 2026
AccessOpen WeightOpen Weight

Which is better: Kimi K2 Thinking or Qwen3.8-27B?

Kimi K2 Thinking and Qwen3.8-27B don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Kimi K2 Thinking shipped 281 days before Qwen3.8-27B, so benchmark comparisons should account for the intervening progress.

Kimi K2 Thinking has 1T parameters, while Qwen3.8-27B has 27B. Context windows are 256k (Kimi K2 Thinking) vs 262k (Qwen3.8-27B).

Direct benchmark comparisons are unavailable — Kimi K2 Thinking and Qwen3.8-27B don't publish scores on any of the same benchmarks.

Frequently asked questions

Kimi K2 Thinking was released by Moonshot AI on Nov 6 2025.

Qwen3.8-27B was released by Qwen on Aug 14 2026.

Kimi K2 Thinking has a 256k context window; Qwen3.8-27B has 262k.

Other comparisons