Compare AI models

Specifications
Parameters
—320B
Context window
—1M
API pricing
Input price
$2.00—
Output price
$12.00—
Cached input price
$0.20—
Cheapest input
$1.00Google$0.04Relace
Cheapest output
$6.00Google$0.23StreamLake
Benchmarks
DeepSWE 1.1
12%63.4%
Terminal-Bench 2.1
70.3%84.3%
Humanity's Last Exam · with tools
51.4%55.3%
GDPval-AA v2
9651773
Overview
CompanyGoogleZ.ai
Release dateFeb 19 2026Aug 26 2026
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

GLM-5.3-Flash leads Gemini 3.1 Pro on 4 of the 4 benchmarks they both report (DeepSWE 1.1, Terminal-Bench 2.1, Humanity's Last Exam, GDPval-AA v2). Only Gemini 3.1 Pro has a verified first-party API price: $2.00 per million input tokens and $12.00 per million output tokens. No pay-as-you-go API rate is tracked for GLM-5.3-Flash. Gemini 3.1 Pro shipped 188 days before GLM-5.3-Flash, so benchmark comparisons should account for the intervening progress.

Gemini 3.1 Pro is closed, while GLM-5.3-Flash is open weight.

On DeepSWE 1.1, GLM-5.3-Flash leads at 63.4% vs Gemini 3.1 Pro at 12%. On Terminal-Bench 2.1, GLM-5.3-Flash leads at 84.3% vs Gemini 3.1 Pro at 70.3%. On Humanity's Last Exam · with tools, GLM-5.3-Flash leads at 55.3% vs Gemini 3.1 Pro at 51.4%. On GDPval-AA v2, GLM-5.3-Flash leads at 1773 vs Gemini 3.1 Pro at 965.