Compare AI models

Specifications
Parameters
—125B
Context window
400k262k
API pricing
Input price
$0.75—
Output price
$4.50—
Cached input price
$0.075—
Cheapest input
$0.375OpenAI—
Cheapest output
$2.25OpenAI—
Benchmarks
Humanity's Last Exam · with tools
41.5%35.9%
Overview
CompanyOpenAIQwen
Release dateMar 17 2026Aug 26 2026
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

GPT-5.4 mini leads Qwen3.8-Flash-Next on 1 of the 1 benchmark they both report (Humanity's Last Exam). Only GPT-5.4 mini has a verified first-party API price: $0.75 per million input tokens and $4.50 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. GPT-5.4 mini shipped 162 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

Context windows are 400k (GPT-5.4 mini) vs 262k (Qwen3.8-Flash-Next). GPT-5.4 mini is closed, while Qwen3.8-Flash-Next is open weight.

On Humanity's Last Exam · with tools, GPT-5.4 mini leads at 41.5% vs Qwen3.8-Flash-Next at 35.9%.