Compare AI models

Specifications
Parameters
1T125B
Context window
256k262k
API pricing
Input price
$0.95—
Output price
$4.00—
Cached input price
$0.16—
Cheapest input
$0.465Inceptron—
Cheapest output
$2.40DigitalOcean—
Benchmarks
LiveCodeBench
89.6%91.9%
Humanity's Last Exam · with tools
34.7%35.9%
GPQA Diamond
90.5%91.7%
Overview
CompanyMoonshot AIQwen
Release dateApr 21 2026Aug 26 2026
AccessOpen WeightOpen Weight
Model detailsView modelView model

Frequently asked questions

Qwen3.8-Flash-Next leads Kimi K2.6 on 3 of the 3 benchmarks they both report (LiveCodeBench, Humanity's Last Exam, GPQA Diamond). Only Kimi K2.6 has a verified first-party API price: $0.95 per million input tokens and $4.00 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. Kimi K2.6 shipped 127 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

Kimi K2.6 has 1T parameters, while Qwen3.8-Flash-Next has 125B. Context windows are 256k (Kimi K2.6) vs 262k (Qwen3.8-Flash-Next).

On LiveCodeBench, Qwen3.8-Flash-Next leads at 91.9% vs Kimi K2.6 at 89.6%. On Humanity's Last Exam · with tools, Qwen3.8-Flash-Next leads at 35.9% vs Kimi K2.6 at 34.7%. On GPQA Diamond, Qwen3.8-Flash-Next leads at 91.7% vs Kimi K2.6 at 90.5%.