Compare AI models

Specifications
Parameters
2.8T125B
Context window
1M262k
API pricing
Input price
$3.00—
Output price
$15.00—
Cached input price
$0.30—
Cheapest input
$0.33Wafer—
Cheapest output
$12.75Makora—
Benchmarks
DeepSWE 1.1
69%58.7%
JobBench
52.9%55.7%
Toolathlon-Verified
73.2%73.5%
Humanity's Last Exam · no tools
43.5%35.9%
Humanity's Last Exam · with tools
56%35.9%
GPQA Diamond
93.5%91.7%
CharXiv Reasoning
84.8%84.6%
Overview
CompanyMoonshot AIQwen
Release dateJul 16 2026Aug 26 2026
AccessOpen WeightOpen Weight
Model detailsView modelView model

Frequently asked questions

Kimi K3 leads Qwen3.8-Flash-Next on 5 of the 7 benchmarks they both report. Only Kimi K3 has a verified first-party API price: $3.00 per million input tokens and $15.00 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. Kimi K3 shipped 41 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

Kimi K3 has 2.8T parameters, while Qwen3.8-Flash-Next has 125B. Context windows are 1M (Kimi K3) vs 262k (Qwen3.8-Flash-Next).

On DeepSWE 1.1, Kimi K3 leads at 69% vs Qwen3.8-Flash-Next at 58.7%. On JobBench, Qwen3.8-Flash-Next leads at 55.7% vs Kimi K3 at 52.9%. On Toolathlon-Verified, Qwen3.8-Flash-Next leads at 73.5% vs Kimi K3 at 73.2%. On Humanity's Last Exam · no tools, Kimi K3 leads at 43.5% vs Qwen3.8-Flash-Next at 35.9%. On Humanity's Last Exam · with tools, Kimi K3 leads at 56% vs Qwen3.8-Flash-Next at 35.9%. On GPQA Diamond, Kimi K3 leads at 93.5% vs Qwen3.8-Flash-Next at 91.7%. On CharXiv Reasoning, Kimi K3 leads at 84.8% vs Qwen3.8-Flash-Next at 84.6%.