Compare AI models

Specifications
Parameters
2.8T320B
Context window
1M1M
API pricing
Input price
$3.00—
Output price
$15.00—
Cached input price
$0.30—
Cheapest input
$0.33Wafer$0.04Relace
Cheapest output
$12.75Makora$0.23StreamLake
Benchmarks
DeepSWE 1.1
69%63.4%
Terminal-Bench 2.1
88.3%84.3%
Toolathlon-Verified
73.2%78.4%
Humanity's Last Exam · with tools
56%55.3%
GDPval-AA v2
16681773
threejseval
15151378
Overview
CompanyMoonshot AIZ.ai
Release dateJul 16 2026Aug 26 2026
AccessOpen WeightOpen Weight
Model detailsView modelView model

Frequently asked questions

Kimi K3 leads GLM-5.3-Flash on 4 of the 6 benchmarks they both report. Only Kimi K3 has a verified first-party API price: $3.00 per million input tokens and $15.00 per million output tokens. No pay-as-you-go API rate is tracked for GLM-5.3-Flash. Kimi K3 shipped 41 days before GLM-5.3-Flash, so benchmark comparisons should account for the intervening progress.

Kimi K3 has 2.8T parameters, while GLM-5.3-Flash has 320B. Context windows are 1M (Kimi K3) vs 1M (GLM-5.3-Flash).

On DeepSWE 1.1, Kimi K3 leads at 69% vs GLM-5.3-Flash at 63.4%. On Terminal-Bench 2.1, Kimi K3 leads at 88.3% vs GLM-5.3-Flash at 84.3%. On Toolathlon-Verified, GLM-5.3-Flash leads at 78.4% vs Kimi K3 at 73.2%. On Humanity's Last Exam · with tools, Kimi K3 leads at 56% vs GLM-5.3-Flash at 55.3%. On GDPval-AA v2, GLM-5.3-Flash leads at 1773 vs Kimi K3 at 1668. On threejseval, Kimi K3 leads at 1515 vs GLM-5.3-Flash at 1378.