Qwen3.8-Flash-NextvsGLM-4.5

Qwen3.8-Flash-Next
GLM-4.5
Specifications
Parameters
125B355B
Context window
262k128k
API pricing
Input price
$0.60
Output price
$2.20
Cached input price
$0.11
Benchmarks
GPQA Diamond
91.7%79.1%
Benchmarks
BullshitBench v2
8%
SWE-Bench Pro
62.5%
SWE-Bench Verified
64.2%
SWE-Bench Multilingual
81%
DeepSWE 1.1
58.7%
NL2Repo-Bench
48.1%
LiveCodeBench
91.9%
JobBench
55.7%
CoWorkBench
73.9%
Toolathlon-Verified
73.5%
Humanity's Last Exam · no tools
35.9%
IFBench
81.3%
OSWorld 2.0
19.4%
Agent's Last Exam · pass@1
24.3%
Agent's Last Exam · score
51.2%
CharXiv Reasoning
84.6%
LVBench
76.6%
Overview
CompanyQwenZ.ai
Release dateAug 26 2026Jul 28 2025
AccessOpen WeightOpen Weight

Other comparisons

Qwen3.8-Flash-NextvsClaude Opus 5GLM-4.5vsClaude Opus 5Qwen3.8-Flash-NextvsGPT-5.6 SolGLM-4.5vsGPT-5.6 SolQwen3.8-Flash-NextvsGemini 3.7 FlashGLM-4.5vsGemini 3.7 FlashQwen3.8-Flash-NextvsMuse GlimmerGLM-4.5vsMuse GlimmerQwen3.8-Flash-NextvsGrok 4.6GLM-4.5vsGrok 4.6Qwen3.8-Flash-NextvsDeepSeek-V4-Pro-0813GLM-4.5vsDeepSeek-V4-Pro-0813

Frequently asked questions

Qwen3.8-Flash-Next leads GLM-4.5 on 1 of the 1 benchmark they both report (GPQA Diamond). Only GLM-4.5 has a verified first-party API price: $0.60 per million input tokens and $2.20 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. GLM-4.5 shipped 394 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

Qwen3.8-Flash-Next has 125B parameters, while GLM-4.5 has 355B. Context windows are 262k (Qwen3.8-Flash-Next) vs 128k (GLM-4.5).

On GPQA Diamond, Qwen3.8-Flash-Next leads at 91.7% vs GLM-4.5 at 79.1%.

Qwen3.8-Flash-Next was released by Qwen on Aug 26 2026.

GLM-4.5 was released by Z.ai on Jul 28 2025.

Qwen3.8-Flash-Next leads on GPQA Diamond — Qwen3.8-Flash-Next 91.7% vs GLM-4.5 79.1%.

Only GLM-4.5 has a verified first-party API price: $0.60 per million input tokens and $2.20 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. Rates are pay-as-you-go API prices verified on August 18, 2026.

Qwen3.8-Flash-Next has a 262k context window; GLM-4.5 has 128k.