Gemini 3.8 FlashvsGLM-4.6

Gemini 3.8 Flash
GLM-4.6
Specifications
Parameters
355B
Context window
1M200k
API pricing
Input price
$0.60
Output price
$2.20
Cached input price
$0.11
Cheapest input
$0.43Venice
Cheapest output
$1.75Venice
Benchmarks
SWE-Bench Verified
68%
DeepSWE 1.1
71%
Terminal-Bench 4.0
19.1%
Terminal-Bench 2.1
89.4%
Humanity's Last Exam (Verified)
54.9%
BioMysteryBench · hard
56.5%
BioMysteryBench · human solved
88.8%
LAB-Bench 2
86.2%
OSWorld 2.0
59%
Finance Agent v2
61.4%
Harvey's Legal Agent Benchmark
10%
GDPval-AA v2
1545
CharXiv Reasoning
86.2%
GDP.PDF
35%
LVBench
87.1%
LVBench · agentic
87.8%
Overview
CompanyGoogleZ.ai
Release dateSep 2 2026Sep 30 2025
AccessProprietaryOpen Weight

Other comparisons

Gemini 3.8 FlashvsClaude Fable 5.1GLM-4.6vsClaude Fable 5.1Gemini 3.8 FlashvsGPT-5.6 SolGLM-4.6vsGPT-5.6 SolGemini 3.8 FlashvsMuse GlimmerGLM-4.6vsMuse GlimmerGemini 3.8 FlashvsGrok 4.6GLM-4.6vsGrok 4.6Gemini 3.8 FlashvsDeepSeek-V4-Pro-0813GLM-4.6vsDeepSeek-V4-Pro-0813Gemini 3.8 FlashvsMistral Medium 3.5GLM-4.6vsMistral Medium 3.5

Frequently asked questions

Gemini 3.8 Flash and GLM-4.6 don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Only GLM-4.6 has a verified first-party API price: $0.60 per million input tokens and $2.20 per million output tokens. No pay-as-you-go API rate is tracked for Gemini 3.8 Flash. GLM-4.6 shipped 337 days before Gemini 3.8 Flash, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Gemini 3.8 Flash) vs 200k (GLM-4.6). Gemini 3.8 Flash is proprietary, while GLM-4.6 is open weight.

Direct benchmark comparisons are unavailable — Gemini 3.8 Flash and GLM-4.6 don't publish scores on any of the same benchmarks.

Gemini 3.8 Flash was released by Google on Sep 2 2026.

GLM-4.6 was released by Z.ai on Sep 30 2025.

Only GLM-4.6 has a verified first-party API price: $0.60 per million input tokens and $2.20 per million output tokens. No pay-as-you-go API rate is tracked for Gemini 3.8 Flash. Rates are pay-as-you-go API prices verified on August 18, 2026.

Gemini 3.8 Flash has a 1M context window; GLM-4.6 has 200k.

Gemini 3.8 Flash is a proprietary model released by Google. GLM-4.6 is an open weight model released by Z.ai.