Compare AI models

Specifications
Parameters
125B744B
Context window
262k—
API pricing
Input price
—$1.00
Output price
—$3.20
Cached input price
—$0.20
Cheapest input
—$0.60GMICloud
Cheapest output
—$1.92GMICloud
Benchmarks
SWE-Bench Multilingual
81%73.3%
Humanity's Last Exam · with tools
35.9%50.4%
GPQA Diamond
91.7%86%
Overview
CompanyQwenZ.ai
Release dateAug 26 2026Feb 12 2026
AccessOpen WeightOpen Weight
Model detailsView modelView model

Frequently asked questions

Qwen3.8-Flash-Next leads GLM-5 on 2 of the 3 benchmarks they both report (SWE-Bench Multilingual, Humanity's Last Exam, GPQA Diamond). Only GLM-5 has a verified first-party API price: $1.00 per million input tokens and $3.20 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. GLM-5 shipped 195 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

Qwen3.8-Flash-Next has 125B parameters, while GLM-5 has 744B.

On SWE-Bench Multilingual, Qwen3.8-Flash-Next leads at 81% vs GLM-5 at 73.3%. On Humanity's Last Exam · with tools, GLM-5 leads at 50.4% vs Qwen3.8-Flash-Next at 35.9%. On GPQA Diamond, Qwen3.8-Flash-Next leads at 91.7% vs GLM-5 at 86%.