Compare AI models

Specifications
Parameters
—320B
Context window
—1M
API pricing
Input price
$1.50—
Output price
$9.00—
Cached input price
$0.15—
Cheapest input
$0.75Google$0.04Relace
Cheapest output
$4.50Google$0.23StreamLake
Benchmarks
DeepSWE 1.1
37%63.4%
Terminal-Bench 2.1
76.2%84.3%
Humanity's Last Exam · with tools
40.2%55.3%
GDPval-AA v2
13491773
Overview
CompanyGoogleZ.ai
Release dateMay 19 2026Aug 26 2026
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

GLM-5.3-Flash leads Gemini 3.5 Flash on 4 of the 4 benchmarks they both report (DeepSWE 1.1, Terminal-Bench 2.1, Humanity's Last Exam, GDPval-AA v2). Only Gemini 3.5 Flash has a verified first-party API price: $1.50 per million input tokens and $9.00 per million output tokens. No pay-as-you-go API rate is tracked for GLM-5.3-Flash. Gemini 3.5 Flash shipped 99 days before GLM-5.3-Flash, so benchmark comparisons should account for the intervening progress.

Gemini 3.5 Flash is closed, while GLM-5.3-Flash is open weight.

On DeepSWE 1.1, GLM-5.3-Flash leads at 63.4% vs Gemini 3.5 Flash at 37%. On Terminal-Bench 2.1, GLM-5.3-Flash leads at 84.3% vs Gemini 3.5 Flash at 76.2%. On Humanity's Last Exam · with tools, GLM-5.3-Flash leads at 55.3% vs Gemini 3.5 Flash at 40.2%. On GDPval-AA v2, GLM-5.3-Flash leads at 1773 vs Gemini 3.5 Flash at 1349.