Claude Sonnet 4.5vsQwen3.5

Claude Sonnet 4.5
Qwen3.5
Specifications
Parameters
397B
Context window
1M
Benchmarks
Nonsense detection
BullshitBench v2
79%
Coding
SWE-Bench Verified
77.2%Best
76.4%
Multilingual coding
SWE-Bench Multilingual
69.3%
Next.js coding
Next.js Evals
50%
Agentic terminal coding
Terminal-Bench 2.0
52.5%
Web browsing
BrowseComp
69%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
28.7%
Science
GPQA Diamond
83.4%
88.4%Best
Agentic computer use
OSWorld-Verified
62.2%
Chart reasoning
CharXiv Reasoning
80.8%
Multimodal reasoning
MMMU-Pro
79%
Multimodal
MMMU
68%
85%Best
Overview
CompanyAnthropicQwen
Release dateSep 29 2025Feb 16 2026
AccessProprietaryOpen Weight

Which is better: Claude Sonnet 4.5 or Qwen3.5?

Qwen3.5 leads Claude Sonnet 4.5 on 2 of the 3 benchmarks they both report (SWE-Bench Verified, GPQA Diamond, MMMU). Claude Sonnet 4.5 shipped 140 days before Qwen3.5, so benchmark comparisons should account for the intervening progress.

Claude Sonnet 4.5 is proprietary, while Qwen3.5 is open weight.

On SWE-Bench Verified, Claude Sonnet 4.5 leads at 77.2% vs Qwen3.5 at 76.4%. On GPQA Diamond, Qwen3.5 leads at 88.4% vs Claude Sonnet 4.5 at 83.4%. On MMMU, Qwen3.5 leads at 85% vs Claude Sonnet 4.5 at 68%.

Frequently asked questions

Claude Sonnet 4.5 was released by Anthropic on Sep 29 2025.

Qwen3.5 was released by Qwen on Feb 16 2026.

Claude Sonnet 4.5 leads on SWE-Bench Verified — Claude Sonnet 4.5 77.2% vs Qwen3.5 76.4%.

Qwen3.5 leads on GPQA Diamond — Claude Sonnet 4.5 83.4% vs Qwen3.5 88.4%.

Claude Sonnet 4.5 is a proprietary model released by Anthropic. Qwen3.5 is an open weight model released by Qwen.

Other comparisons