Compare AI models

Specifications
Parameters
—320B
Context window
—1M
API pricing
Input price
$0.22—
Output price
$0.66—
Cached input price
$0.007—
Cheapest input
$0.0099Relace$0.04Relace
Cheapest output
$0.18DeepInfra$0.23StreamLake
Benchmarks
Terminal-Bench 2.1
82.7%84.3%
Toolathlon-Verified
70.3%78.4%
Humanity's Last Exam · with tools
34.8%55.3%
AutomationBench
25.1%48.8%
Overview
CompanyDeepSeekZ.ai
Release dateJul 31 2026Aug 26 2026
AccessOpen WeightOpen Weight
Model detailsView modelView model

Frequently asked questions

GLM-5.3-Flash leads DeepSeek-V4-Flash-0731 on 4 of the 4 benchmarks they both report (Terminal-Bench 2.1, Toolathlon-Verified, Humanity's Last Exam, AutomationBench). Only DeepSeek-V4-Flash-0731 has a verified first-party API price: $0.22 per million input tokens and $0.66 per million output tokens. No pay-as-you-go API rate is tracked for GLM-5.3-Flash. DeepSeek-V4-Flash-0731 shipped 26 days before GLM-5.3-Flash, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On Terminal-Bench 2.1, GLM-5.3-Flash leads at 84.3% vs DeepSeek-V4-Flash-0731 at 82.7%. On Toolathlon-Verified, GLM-5.3-Flash leads at 78.4% vs DeepSeek-V4-Flash-0731 at 70.3%. On Humanity's Last Exam · with tools, GLM-5.3-Flash leads at 55.3% vs DeepSeek-V4-Flash-0731 at 34.8%. On AutomationBench, GLM-5.3-Flash leads at 48.8% vs DeepSeek-V4-Flash-0731 at 25.1%.