Gemini 3.8 FlashvsGLM-5.1

Gemini 3.8 Flash
GLM-5.1
Specifications
Parameters
744B
Context window
1M200k
API pricing
Input price
$1.40
Output price
$4.40
Cached input price
$0.26
Cheapest input
$0.91GMICloud
Cheapest output
$2.86GMICloud
Benchmarks
BullshitBench v2
22%
DeepSWE 1.1
71%
Next.js Evals
69%
Terminal-Bench 4.0
19.1%
Terminal-Bench 2.1
89.4%
Terminal-Bench 2.0
63.5%
BrowseComp
68%
Humanity's Last Exam · with tools
52.3%
Humanity's Last Exam (Verified)
54.9%
BioMysteryBench · hard
56.5%
BioMysteryBench · human solved
88.8%
LAB-Bench 2
86.2%
GPQA Diamond
86.2%
OSWorld 2.0
59%
Finance Agent v2
61.4%
Harvey's Legal Agent Benchmark
10%
GDPval-AA v2
1545
CharXiv Reasoning
86.2%
GDP.PDF
35%
LVBench
87.1%
LVBench · agentic
87.8%
Overview
CompanyGoogleZ.ai
Release dateSep 2 2026Apr 7 2026
AccessProprietaryOpen Weight

Other comparisons

Gemini 3.8 FlashvsClaude Fable 5.1GLM-5.1vsClaude Fable 5.1Gemini 3.8 FlashvsGPT-5.6 SolGLM-5.1vsGPT-5.6 SolGemini 3.8 FlashvsMuse GlimmerGLM-5.1vsMuse GlimmerGemini 3.8 FlashvsGrok 4.6GLM-5.1vsGrok 4.6Gemini 3.8 FlashvsDeepSeek-V4-Pro-0813GLM-5.1vsDeepSeek-V4-Pro-0813Gemini 3.8 FlashvsMistral Medium 3.5GLM-5.1vsMistral Medium 3.5

Frequently asked questions

Gemini 3.8 Flash and GLM-5.1 don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Only GLM-5.1 has a verified first-party API price: $1.40 per million input tokens and $4.40 per million output tokens. No pay-as-you-go API rate is tracked for Gemini 3.8 Flash. GLM-5.1 shipped 148 days before Gemini 3.8 Flash, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Gemini 3.8 Flash) vs 200k (GLM-5.1). Gemini 3.8 Flash is proprietary, while GLM-5.1 is open weight.

Direct benchmark comparisons are unavailable — Gemini 3.8 Flash and GLM-5.1 don't publish scores on any of the same benchmarks.

Gemini 3.8 Flash was released by Google on Sep 2 2026.

GLM-5.1 was released by Z.ai on Apr 7 2026.

Only GLM-5.1 has a verified first-party API price: $1.40 per million input tokens and $4.40 per million output tokens. No pay-as-you-go API rate is tracked for Gemini 3.8 Flash. Rates are pay-as-you-go API prices verified on August 18, 2026.

Gemini 3.8 Flash has a 1M context window; GLM-5.1 has 200k.

Gemini 3.8 Flash is a proprietary model released by Google. GLM-5.1 is an open weight model released by Z.ai.