Qwen3.8-Flash-NextvsGLM-5.1

Qwen3.8-Flash-Next
GLM-5.1
Specifications
Parameters
125B744B
Context window
262k200k
API pricing
Input price
$1.40
Output price
$4.40
Cached input price
$0.26
Cheapest input
$0.91GMICloud
Cheapest output
$2.86GMICloud
Benchmarks
GPQA Diamond
91.7%86.2%
Benchmarks
BullshitBench v2
22%
SWE-Bench Pro
62.5%
SWE-Bench Multilingual
81%
DeepSWE 1.1
58.7%
Next.js Evals
75%
NL2Repo-Bench
48.1%
LiveCodeBench
91.9%
Terminal-Bench 2.0
63.5%
JobBench
55.7%
CoWorkBench
73.9%
Toolathlon-Verified
73.5%
BrowseComp
68%
Humanity's Last Exam · no tools
35.9%
Humanity's Last Exam · with tools
52.3%
IFBench
81.3%
OSWorld 2.0
19.4%
Agent's Last Exam · pass@1
24.3%
Agent's Last Exam · score
51.2%
CharXiv Reasoning
84.6%
LVBench
76.6%
Overview
CompanyQwenZ.ai
Release dateAug 26 2026Apr 7 2026
AccessOpen WeightOpen Weight

Other comparisons

Qwen3.8-Flash-NextvsClaude Opus 5GLM-5.1vsClaude Opus 5Qwen3.8-Flash-NextvsGPT-5.6 SolGLM-5.1vsGPT-5.6 SolQwen3.8-Flash-NextvsGemini 3.7 FlashGLM-5.1vsGemini 3.7 FlashQwen3.8-Flash-NextvsMuse GlimmerGLM-5.1vsMuse GlimmerQwen3.8-Flash-NextvsGrok 4.6GLM-5.1vsGrok 4.6Qwen3.8-Flash-NextvsDeepSeek-V4-Pro-0813GLM-5.1vsDeepSeek-V4-Pro-0813

Frequently asked questions

Qwen3.8-Flash-Next leads GLM-5.1 on 1 of the 1 benchmark they both report (GPQA Diamond). Only GLM-5.1 has a verified first-party API price: $1.40 per million input tokens and $4.40 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. GLM-5.1 shipped 141 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

Qwen3.8-Flash-Next has 125B parameters, while GLM-5.1 has 744B. Context windows are 262k (Qwen3.8-Flash-Next) vs 200k (GLM-5.1).

On GPQA Diamond, Qwen3.8-Flash-Next leads at 91.7% vs GLM-5.1 at 86.2%.

Qwen3.8-Flash-Next was released by Qwen on Aug 26 2026.

GLM-5.1 was released by Z.ai on Apr 7 2026.

Qwen3.8-Flash-Next leads on GPQA Diamond — Qwen3.8-Flash-Next 91.7% vs GLM-5.1 86.2%.

Only GLM-5.1 has a verified first-party API price: $1.40 per million input tokens and $4.40 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. Rates are pay-as-you-go API prices verified on August 18, 2026.

Qwen3.8-Flash-Next has a 262k context window; GLM-5.1 has 200k.