Compare AI models

Specifications
Parameters
—320B
Context window
1M1M
API pricing
Input price
$0.75—
Output price
$3.75—
Cached input price
$0.075—
Cheapest input
$0.375Google$0.04Relace
Cheapest output
$1.875Google$0.23StreamLake
Benchmarks
DeepSWE 1.1
65.3%63.4%
Terminal-Bench 2.1
85.8%84.3%
Agent's Last Exam · pass@1
26.3%26.3%
AutomationBench
30.4%48.8%
GDPval-AA v2
15251773
threejseval
14711378
Overview
CompanyGoogleZ.ai
Release dateAug 13 2026Aug 26 2026
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

Gemini 3.7 Flash leads GLM-5.3-Flash on 3 of the 6 benchmarks they both report. Only Gemini 3.7 Flash has a verified first-party API price: $0.75 per million input tokens and $3.75 per million output tokens. No pay-as-you-go API rate is tracked for GLM-5.3-Flash. Gemini 3.7 Flash shipped 13 days before GLM-5.3-Flash, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Gemini 3.7 Flash) vs 1M (GLM-5.3-Flash). Gemini 3.7 Flash is closed, while GLM-5.3-Flash is open weight.

On DeepSWE 1.1, Gemini 3.7 Flash leads at 65.3% vs GLM-5.3-Flash at 63.4%. On Terminal-Bench 2.1, Gemini 3.7 Flash leads at 85.8% vs GLM-5.3-Flash at 84.3%. On Agent's Last Exam · pass@1, both models score 26.3%. On AutomationBench, GLM-5.3-Flash leads at 48.8% vs Gemini 3.7 Flash at 30.4%. On GDPval-AA v2, GLM-5.3-Flash leads at 1773 vs Gemini 3.7 Flash at 1525. On threejseval, Gemini 3.7 Flash leads at 1471 vs GLM-5.3-Flash at 1378.