Claude Opus 4.8vsQwen3.8-Flash-Next

Claude Opus 4.8
Qwen3.8-Flash-Next
Specifications
Parameters
125B
Context window
1M262k
API pricing
Input price
$5.00
Output price
$25.00
Cached input price
$0.50
Cheapest input
$5.00Amazon Bedrock
Cheapest output
$25.00Amazon Bedrock
Benchmarks
SWE-Bench Pro
69.2%62.5%
SWE-Bench Multilingual
84.4%81%
Humanity's Last Exam · no tools
49.8%35.9%
Benchmarks
BullshitBench v2
95%
Gray Swan IPI · k = 1
0.5%
Gray Swan IPI · k = 10
4.1%
Gray Swan IPI · k = 15
5.5%
SWE-Bench Verified
88.6%
DeepSWE 1.1
58.7%
DeepSWE 1.0
55.8%
Next.js Evals
88%
NL2Repo-Bench
48.1%
LiveCodeBench
91.9%
Frontier-Bench v0.1
21.1%
Terminal-Bench 2.1
74.6%
JobBench
55.7%
CoWorkBench
73.9%
Toolathlon-Verified
73.5%
BU Bench
74%
BrowseComp
84.3%
Humanity's Last Exam · with tools
57.9%
GPQA Diamond
91.7%
IFBench
81.3%
OSWorld 2.0
19.4%
OSWorld-Verified
83.4%
Agent's Last Exam · pass@1
24.3%
Agent's Last Exam · score
51.2%
Finance Agent v2
53.9%
Harvey's Legal Agent Benchmark
9.58%
TaxEval v2
75.63%
MedScribe
85.75%
GDPval-AA
1890
GDPval-AA v2
1600
CharXiv Reasoning
84.6%
LVBench
76.6%
Overview
CompanyAnthropicQwen
Release dateMay 28 2026Aug 26 2026
AccessProprietaryOpen Weight

Other comparisons

Claude Opus 4.8vsGPT-5.6 SolQwen3.8-Flash-NextvsGPT-5.6 SolClaude Opus 4.8vsGemini 3.7 FlashQwen3.8-Flash-NextvsGemini 3.7 FlashClaude Opus 4.8vsMuse GlimmerQwen3.8-Flash-NextvsMuse GlimmerClaude Opus 4.8vsGrok 4.6Qwen3.8-Flash-NextvsGrok 4.6Claude Opus 4.8vsDeepSeek-V4-Pro-0813Qwen3.8-Flash-NextvsDeepSeek-V4-Pro-0813Claude Opus 4.8vsMistral Medium 3.5Qwen3.8-Flash-NextvsMistral Medium 3.5

Frequently asked questions

Claude Opus 4.8 leads Qwen3.8-Flash-Next on 3 of the 3 benchmarks they both report (SWE-Bench Pro, SWE-Bench Multilingual, Humanity's Last Exam). Only Claude Opus 4.8 has a verified first-party API price: $5.00 per million input tokens and $25.00 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. Claude Opus 4.8 shipped 90 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Opus 4.8) vs 262k (Qwen3.8-Flash-Next). Claude Opus 4.8 is proprietary, while Qwen3.8-Flash-Next is open weight.

On SWE-Bench Pro, Claude Opus 4.8 leads at 69.2% vs Qwen3.8-Flash-Next at 62.5%. On SWE-Bench Multilingual, Claude Opus 4.8 leads at 84.4% vs Qwen3.8-Flash-Next at 81%. On Humanity's Last Exam · no tools, Claude Opus 4.8 leads at 49.8% vs Qwen3.8-Flash-Next at 35.9%.

Claude Opus 4.8 was released by Anthropic on May 28 2026.

Qwen3.8-Flash-Next was released by Qwen on Aug 26 2026.

Claude Opus 4.8 leads on SWE-Bench Pro — Claude Opus 4.8 69.2% vs Qwen3.8-Flash-Next 62.5%.

Claude Opus 4.8 leads on Humanity's Last Exam · no tools — Claude Opus 4.8 49.8% vs Qwen3.8-Flash-Next 35.9%.

Only Claude Opus 4.8 has a verified first-party API price: $5.00 per million input tokens and $25.00 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. Rates are pay-as-you-go API prices verified on August 18, 2026.

Claude Opus 4.8 has a 1M context window; Qwen3.8-Flash-Next has 262k.

Claude Opus 4.8 is a proprietary model released by Anthropic. Qwen3.8-Flash-Next is an open weight model released by Qwen.