Compare AI models

Specifications
Parameters
—125B
Context window
128k262k
API pricing
Input price
$2.50—
Output price
$10.00—
Cached input price
$1.25—
Cheapest input
$2.50Azure—
Cheapest output
$10.00Azure—
Benchmarks
GPQA Diamond
49.9%91.7%
Overview
CompanyOpenAIQwen
Release dateMay 13 2024Aug 26 2026
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

Qwen3.8-Flash-Next leads GPT-4o on 1 of the 1 benchmark they both report (GPQA Diamond). Only GPT-4o has a verified first-party API price: $2.50 per million input tokens and $10.00 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. GPT-4o shipped 835 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

Context windows are 128k (GPT-4o) vs 262k (Qwen3.8-Flash-Next). GPT-4o is closed, while Qwen3.8-Flash-Next is open weight.

On GPQA Diamond, Qwen3.8-Flash-Next leads at 91.7% vs GPT-4o at 49.9%.