Compare AI models

Specifications
Parameters
—320B
Context window
—1M
API pricing
Input price
$0.30—
Output price
$2.50—
Cached input price
$0.03—
Cheapest input
$0.15Google$0.04Relace
Cheapest output
$1.25Google$0.23StreamLake
Benchmarks
Terminal-Bench 2.1
54%84.3%
GDPval-AA v2
11401773
Overview
CompanyGoogleZ.ai
Release dateJul 21 2026Aug 26 2026
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

GLM-5.3-Flash leads Gemini 3.5 Flash-Lite on 2 of the 2 benchmarks they both report (Terminal-Bench 2.1, GDPval-AA v2). Only Gemini 3.5 Flash-Lite has a verified first-party API price: $0.30 per million input tokens and $2.50 per million output tokens. No pay-as-you-go API rate is tracked for GLM-5.3-Flash. Gemini 3.5 Flash-Lite shipped 36 days before GLM-5.3-Flash, so benchmark comparisons should account for the intervening progress.

Gemini 3.5 Flash-Lite is closed, while GLM-5.3-Flash is open weight.

On Terminal-Bench 2.1, GLM-5.3-Flash leads at 84.3% vs Gemini 3.5 Flash-Lite at 54%. On GDPval-AA v2, GLM-5.3-Flash leads at 1773 vs Gemini 3.5 Flash-Lite at 1140.