Compare AI models

Specifications
Parameters
—125B
Context window
128k262k
API pricing
Input price
$0.15—
Output price
$0.60—
Cached input price
$0.075—
Cheapest input
$0.15Azure—
Cheapest output
$0.60Azure—
Benchmarks
GPQA Diamond
40.2%91.7%
Overview
CompanyOpenAIQwen
Release dateJul 18 2024Aug 26 2026
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

Qwen3.8-Flash-Next leads GPT-4o mini on 1 of the 1 benchmark they both report (GPQA Diamond). Only GPT-4o mini has a verified first-party API price: $0.15 per million input tokens and $0.60 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. GPT-4o mini shipped 769 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

Context windows are 128k (GPT-4o mini) vs 262k (Qwen3.8-Flash-Next). GPT-4o mini is closed, while Qwen3.8-Flash-Next is open weight.

On GPQA Diamond, Qwen3.8-Flash-Next leads at 91.7% vs GPT-4o mini at 40.2%.