Qwen3.8-Flash-NextvsGLM-4.6

Qwen3.8-Flash-Next
GLM-4.6
Specifications
Parameters
125B355B
Context window
262k200k
API pricing
Input price
$0.60
Output price
$2.20
Cached input price
$0.11
Cheapest input
$0.43Venice
Cheapest output
$1.75Venice
Benchmarks
SWE-Bench Pro
62.5%
SWE-Bench Verified
68%
SWE-Bench Multilingual
81%
DeepSWE 1.1
58.7%
NL2Repo-Bench
48.1%
LiveCodeBench
91.9%
JobBench
55.7%
CoWorkBench
73.9%
Toolathlon-Verified
73.5%
Humanity's Last Exam · no tools
35.9%
GPQA Diamond
91.7%
IFBench
81.3%
OSWorld 2.0
19.4%
Agent's Last Exam · pass@1
24.3%
Agent's Last Exam · score
51.2%
CharXiv Reasoning
84.6%
LVBench
76.6%
Overview
CompanyQwenZ.ai
Release dateAug 26 2026Sep 30 2025
AccessOpen WeightOpen Weight

Other comparisons

Qwen3.8-Flash-NextvsClaude Opus 5GLM-4.6vsClaude Opus 5Qwen3.8-Flash-NextvsGPT-5.6 SolGLM-4.6vsGPT-5.6 SolQwen3.8-Flash-NextvsGemini 3.7 FlashGLM-4.6vsGemini 3.7 FlashQwen3.8-Flash-NextvsMuse GlimmerGLM-4.6vsMuse GlimmerQwen3.8-Flash-NextvsGrok 4.6GLM-4.6vsGrok 4.6Qwen3.8-Flash-NextvsDeepSeek-V4-Pro-0813GLM-4.6vsDeepSeek-V4-Pro-0813

Frequently asked questions

Qwen3.8-Flash-Next and GLM-4.6 don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Only GLM-4.6 has a verified first-party API price: $0.60 per million input tokens and $2.20 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. GLM-4.6 shipped 330 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

Qwen3.8-Flash-Next has 125B parameters, while GLM-4.6 has 355B. Context windows are 262k (Qwen3.8-Flash-Next) vs 200k (GLM-4.6).

Direct benchmark comparisons are unavailable — Qwen3.8-Flash-Next and GLM-4.6 don't publish scores on any of the same benchmarks.

Qwen3.8-Flash-Next was released by Qwen on Aug 26 2026.

GLM-4.6 was released by Z.ai on Sep 30 2025.

Only GLM-4.6 has a verified first-party API price: $0.60 per million input tokens and $2.20 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. Rates are pay-as-you-go API prices verified on August 18, 2026.

Qwen3.8-Flash-Next has a 262k context window; GLM-4.6 has 200k.