Compare AI models

Specifications
Parameters
1T320B
Context window
256k1M
API pricing
Input price
$0.60—
Output price
$3.00—
Cached input price
$0.10—
Cheapest input
$0.50SiliconFlow$0.04Relace
Cheapest output
$2.50SiliconFlow$0.23StreamLake
Benchmarks
Humanity's Last Exam · with tools
30.1%55.3%
Overview
CompanyMoonshot AIZ.ai
Release dateJan 27 2026Aug 26 2026
AccessOpen WeightOpen Weight
Model detailsView modelView model

Frequently asked questions

GLM-5.3-Flash leads Kimi K2.5 on 1 of the 1 benchmark they both report (Humanity's Last Exam). Only Kimi K2.5 has a verified first-party API price: $0.60 per million input tokens and $3.00 per million output tokens. No pay-as-you-go API rate is tracked for GLM-5.3-Flash. Kimi K2.5 shipped 211 days before GLM-5.3-Flash, so benchmark comparisons should account for the intervening progress.

Kimi K2.5 has 1T parameters, while GLM-5.3-Flash has 320B. Context windows are 256k (Kimi K2.5) vs 1M (GLM-5.3-Flash).

On Humanity's Last Exam · with tools, GLM-5.3-Flash leads at 55.3% vs Kimi K2.5 at 30.1%.