Compare AI models

Specifications
Parameters
355B320B
Context window
128k1M
API pricing
Input price
$0.60—
Output price
$2.20—
Cached input price
$0.11—
Cheapest input
$0.60Z.AI$0.04Relace
Cheapest output
$2.20Z.AI$0.23StreamLake

These models have no shared benchmark scores.

Overview
CompanyZ.aiZ.ai
Release dateJul 28 2025Aug 26 2026
AccessOpen WeightOpen Weight
Model detailsView modelView model

Frequently asked questions

GLM-4.5 and GLM-5.3-Flash don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Only GLM-4.5 has a verified first-party API price: $0.60 per million input tokens and $2.20 per million output tokens. No pay-as-you-go API rate is tracked for GLM-5.3-Flash. GLM-4.5 shipped 394 days before GLM-5.3-Flash, so benchmark comparisons should account for the intervening progress.

GLM-4.5 has 355B parameters, while GLM-5.3-Flash has 320B. Context windows are 128k (GLM-4.5) vs 1M (GLM-5.3-Flash).

Direct benchmark comparisons are unavailable — GLM-4.5 and GLM-5.3-Flash don't publish scores on any of the same benchmarks.