Compare AI models

Specifications
Parameters
117B125B
Context window
128k262k
API pricing
Cheapest input
$0.03AkashML—
Cheapest output
$0.15Venice—
Benchmarks
Humanity's Last Exam · no tools
14.9%35.9%
Humanity's Last Exam · with tools
19%35.9%
GPQA Diamond
80.1%91.7%
Overview
CompanyOpenAIQwen
Release dateAug 5 2025Aug 26 2026
AccessOpen WeightOpen Weight
Model detailsView modelView model

Frequently asked questions

Qwen3.8-Flash-Next leads gpt-oss-120b on 3 of the 3 benchmarks they both report (Humanity's Last Exam, GPQA Diamond). gpt-oss-120b shipped 386 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

gpt-oss-120b has 117B parameters, while Qwen3.8-Flash-Next has 125B. Context windows are 128k (gpt-oss-120b) vs 262k (Qwen3.8-Flash-Next).

On Humanity's Last Exam · no tools, Qwen3.8-Flash-Next leads at 35.9% vs gpt-oss-120b at 14.9%. On Humanity's Last Exam · with tools, Qwen3.8-Flash-Next leads at 35.9% vs gpt-oss-120b at 19%. On GPQA Diamond, Qwen3.8-Flash-Next leads at 91.7% vs gpt-oss-120b at 80.1%.