Compare AI models

Specifications
Parameters
—320B
Context window
400k1M
API pricing
Input price
$0.75—
Output price
$4.50—
Cached input price
$0.075—
Cheapest input
$0.375OpenAI$0.04Relace
Cheapest output
$2.25OpenAI$0.23StreamLake
Benchmarks
Humanity's Last Exam · with tools
41.5%55.3%
Overview
CompanyOpenAIZ.ai
Release dateMar 17 2026Aug 26 2026
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

GLM-5.3-Flash leads GPT-5.4 mini on 1 of the 1 benchmark they both report (Humanity's Last Exam). Only GPT-5.4 mini has a verified first-party API price: $0.75 per million input tokens and $4.50 per million output tokens. No pay-as-you-go API rate is tracked for GLM-5.3-Flash. GPT-5.4 mini shipped 162 days before GLM-5.3-Flash, so benchmark comparisons should account for the intervening progress.

Context windows are 400k (GPT-5.4 mini) vs 1M (GLM-5.3-Flash). GPT-5.4 mini is closed, while GLM-5.3-Flash is open weight.

On Humanity's Last Exam · with tools, GLM-5.3-Flash leads at 55.3% vs GPT-5.4 mini at 41.5%.