Compare AI models

Specifications
Parameters
—125B
Context window
—262k
API pricing
Input price
$3.00—
Output price
$15.00—
Cached input price
$0.30—
Cheapest input
$3.00Amazon Bedrock—
Cheapest output
$15.00Amazon Bedrock—
Benchmarks
DeepSWE 1.1
30%58.7%
Humanity's Last Exam · no tools
33.2%35.9%
Humanity's Last Exam · with tools
49%35.9%
GPQA Diamond
89.9%91.7%
CharXiv Reasoning
72.4%84.6%
Overview
CompanyAnthropicQwen
Release dateFeb 17 2026Aug 26 2026
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

Qwen3.8-Flash-Next leads Claude Sonnet 4.6 on 4 of the 5 benchmarks they both report (DeepSWE 1.1, Humanity's Last Exam, GPQA Diamond, CharXiv Reasoning). Only Claude Sonnet 4.6 has a verified first-party API price: $3.00 per million input tokens and $15.00 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. Claude Sonnet 4.6 shipped 190 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

Claude Sonnet 4.6 is closed, while Qwen3.8-Flash-Next is open weight.

On DeepSWE 1.1, Qwen3.8-Flash-Next leads at 58.7% vs Claude Sonnet 4.6 at 30%. On Humanity's Last Exam · no tools, Qwen3.8-Flash-Next leads at 35.9% vs Claude Sonnet 4.6 at 33.2%. On Humanity's Last Exam · with tools, Claude Sonnet 4.6 leads at 49% vs Qwen3.8-Flash-Next at 35.9%. On GPQA Diamond, Qwen3.8-Flash-Next leads at 91.7% vs Claude Sonnet 4.6 at 89.9%. On CharXiv Reasoning, Qwen3.8-Flash-Next leads at 84.6% vs Claude Sonnet 4.6 at 72.4%.