Compare AI models

Specifications
Parameters
21B125B
Context window
128k262k
API pricing
Cheapest input
$0.018Darkbloom—
Cheapest output
$0.09Darkbloom—
Benchmarks
Humanity's Last Exam · no tools
10.9%35.9%
Humanity's Last Exam · with tools
17.3%35.9%
GPQA Diamond
71.5%91.7%
Overview
CompanyOpenAIQwen
Release dateAug 5 2025Aug 26 2026
AccessOpen WeightOpen Weight
Model detailsView modelView model

Frequently asked questions

Qwen3.8-Flash-Next leads gpt-oss-20b on 3 of the 3 benchmarks they both report (Humanity's Last Exam, GPQA Diamond). gpt-oss-20b shipped 386 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

gpt-oss-20b has 21B parameters, while Qwen3.8-Flash-Next has 125B. Context windows are 128k (gpt-oss-20b) vs 262k (Qwen3.8-Flash-Next).

On Humanity's Last Exam · no tools, Qwen3.8-Flash-Next leads at 35.9% vs gpt-oss-20b at 10.9%. On Humanity's Last Exam · with tools, Qwen3.8-Flash-Next leads at 35.9% vs gpt-oss-20b at 17.3%. On GPQA Diamond, Qwen3.8-Flash-Next leads at 91.7% vs gpt-oss-20b at 71.5%.