Kimi K2.6vsQwen3.5

Kimi K2.6
Qwen3.5
Specifications
Parameters
1T
397B
Context window
256k
1M
Benchmarks
Nonsense detection
BullshitBench v2
65%
Coding
SWE-Bench Verified
80.2%Best
76.4%
Multilingual coding
SWE-Bench Multilingual
69.3%
Agentic coding
CursorBench v3.1
47.6%
Next.js coding
Next.js Evals
67%
Competitive coding
LiveCodeBench
89.6%
Agentic terminal coding
Terminal-Bench 2.0
52.5%
Web browsing
BrowseComp
69%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
28.7%
Science
GPQA Diamond
90.5%Best
88.4%
Agentic computer use
OSWorld-Verified
62.2%
Chart reasoning
CharXiv Reasoning
80.8%
Multimodal reasoning
MMMU-Pro
79%
Multimodal
MMMU
85%
Community preference (code)
Arena Elo (Code)
1513
Overview
CompanyMoonshot AIQwen
Release dateApr 21 2026Feb 16 2026
AccessOpen WeightOpen Weight

Which is better: Kimi K2.6 or Qwen3.5?

Kimi K2.6 leads Qwen3.5 on 2 of the 2 benchmarks they both report (SWE-Bench Verified, GPQA Diamond). Qwen3.5 shipped 64 days before Kimi K2.6, so benchmark comparisons should account for the intervening progress.

Kimi K2.6 has 1T parameters, while Qwen3.5 has 397B. Context windows are 256k (Kimi K2.6) vs 1M (Qwen3.5).

On SWE-Bench Verified, Kimi K2.6 leads at 80.2% vs Qwen3.5 at 76.4%. On GPQA Diamond, Kimi K2.6 leads at 90.5% vs Qwen3.5 at 88.4%.

Frequently asked questions

Kimi K2.6 was released by Moonshot AI on Apr 21 2026.

Qwen3.5 was released by Qwen on Feb 16 2026.

Kimi K2.6 leads on SWE-Bench Verified — Kimi K2.6 80.2% vs Qwen3.5 76.4%.

Kimi K2.6 leads on GPQA Diamond — Kimi K2.6 90.5% vs Qwen3.5 88.4%.

Kimi K2.6 has a 256k context window; Qwen3.5 has 1M.

Other comparisons