Claude Opus 4.8vsQwen3

Claude Opus 4.8
Qwen3
Specifications
Parameters
235B
Context window
1M128k
API pricing
Input price
$5.00$0.70
Output price
$25.00$2.80
Cached input price
$0.50
Cheapest input
$5.00Amazon Bedrock$0.0482StreamLake
Cheapest output
$25.00Amazon Bedrock$0.1931StreamLake
Benchmarks
BullshitBench v2
95%
Gray Swan IPI · k = 1
0.5%
Gray Swan IPI · k = 10
4.1%
Gray Swan IPI · k = 15
5.5%
ProgramBench
0%
SWE-Bench Pro
69.2%
SWE-Bench Verified
88.6%
SWE-Bench Multilingual
84.4%
DeepSWE 1.0
55.8%
Next.js Evals
68%
Frontier-Bench v0.1
21.1%
Terminal-Bench 4.0
23.64%
Terminal-Bench 2.1
74.6%
BU Bench
74%
BrowseComp
84.3%
Humanity's Last Exam · no tools
49.8%
Humanity's Last Exam · with tools
57.9%
ARC-AGI-2
72.08%
OSWorld-Verified
83.4%
Finance Agent v2
53.9%
Harvey's Legal Agent Benchmark
9.58%
TaxEval v2
75.63%
MedScribe
85.75%
GDPval-AA
1890
GDPval-AA v2
1600
Overview
CompanyAnthropicQwen
Release dateMay 28 2026Apr 29 2025
AccessProprietaryOpen Weight

Other comparisons

Claude Opus 4.8vsGPT-6 AstraQwen3vsGPT-6 AstraClaude Opus 4.8vsGemini 3.8 FlashQwen3vsGemini 3.8 FlashClaude Opus 4.8vsMuse Spark 1.3Qwen3vsMuse Spark 1.3Claude Opus 4.8vsGrok 4.6Qwen3vsGrok 4.6Claude Opus 4.8vsDeepSeek-V4.1-FlashQwen3vsDeepSeek-V4.1-FlashClaude Opus 4.8vsMistral Medium 3.5Qwen3vsMistral Medium 3.5

Frequently asked questions

Claude Opus 4.8 and Qwen3 don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Qwen3 is cheaper on both input and output: $0.70 vs $5.00 per million input tokens, and $2.80 vs $25.00 per million output tokens. Figures are base-tier rates. Qwen3 shipped 394 days before Claude Opus 4.8, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Opus 4.8) vs 128k (Qwen3). Claude Opus 4.8 is proprietary, while Qwen3 is open weight.

Direct benchmark comparisons are unavailable — Claude Opus 4.8 and Qwen3 don't publish scores on any of the same benchmarks.