Claude 3.7 SonnetvsQwen3.5

Claude 3.7 Sonnet
Qwen3.5
Specifications
Parameters
397B
Context window
1M
Benchmarks
Nonsense detection
BullshitBench v2
49%
Coding
SWE-Bench Verified
62.3%
76.4%Best
Multilingual coding
SWE-Bench Multilingual
69.3%
Agentic terminal coding
Terminal-Bench 2.0
52.5%
Web browsing
BrowseComp
69%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
28.7%
Science
GPQA Diamond
68%
88.4%Best
Agentic computer use
OSWorld-Verified
62.2%
Chart reasoning
CharXiv Reasoning
80.8%
Multimodal reasoning
MMMU-Pro
79%
Multimodal
MMMU
85%
Overview
CompanyAnthropicQwen
Release dateFeb 24 2025Feb 16 2026
AccessProprietaryOpen Weight

Which is better: Claude 3.7 Sonnet or Qwen3.5?

Qwen3.5 leads Claude 3.7 Sonnet on 2 of the 2 benchmarks they both report (SWE-Bench Verified, GPQA Diamond). Claude 3.7 Sonnet shipped 357 days before Qwen3.5, so benchmark comparisons should account for the intervening progress.

Claude 3.7 Sonnet is proprietary, while Qwen3.5 is open weight.

On SWE-Bench Verified, Qwen3.5 leads at 76.4% vs Claude 3.7 Sonnet at 62.3%. On GPQA Diamond, Qwen3.5 leads at 88.4% vs Claude 3.7 Sonnet at 68%.

Frequently asked questions

Claude 3.7 Sonnet was released by Anthropic on Feb 24 2025.

Qwen3.5 was released by Qwen on Feb 16 2026.

Qwen3.5 leads on SWE-Bench Verified — Claude 3.7 Sonnet 62.3% vs Qwen3.5 76.4%.

Qwen3.5 leads on GPQA Diamond — Claude 3.7 Sonnet 68% vs Qwen3.5 88.4%.

Claude 3.7 Sonnet is a proprietary model released by Anthropic. Qwen3.5 is an open weight model released by Qwen.

Other comparisons