Claude 3.5 SonnetvsQwen3.6

Claude 3.5 Sonnet
Qwen3.6
Specifications
Parameters
35B
Context window
256k
Benchmarks
Nonsense detection
BullshitBench v2
45%
Agentic coding
SWE-Bench Pro
49.5%
Coding
SWE-Bench Verified
33.4%
73.4%Best
Multilingual coding
SWE-Bench Multilingual
67.2%
Agentic terminal coding
Terminal-Bench 2.0
51.5%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
21.4%
Science
GPQA Diamond
59.4%
86%Best
Chart reasoning
CharXiv Reasoning
78%
Multimodal reasoning
MMMU-Pro
75.3%
Multimodal
MMMU
81.7%
Overview
CompanyAnthropicQwen
Release dateJun 20 2024Apr 16 2026
AccessProprietaryOpen Weight

Which is better: Claude 3.5 Sonnet or Qwen3.6?

Qwen3.6 leads Claude 3.5 Sonnet on 2 of the 2 benchmarks they both report (SWE-Bench Verified, GPQA Diamond). Claude 3.5 Sonnet shipped 665 days before Qwen3.6, so benchmark comparisons should account for the intervening progress.

Claude 3.5 Sonnet is proprietary, while Qwen3.6 is open weight.

On SWE-Bench Verified, Qwen3.6 leads at 73.4% vs Claude 3.5 Sonnet at 33.4%. On GPQA Diamond, Qwen3.6 leads at 86% vs Claude 3.5 Sonnet at 59.4%.

Frequently asked questions

Claude 3.5 Sonnet was released by Anthropic on Jun 20 2024.

Qwen3.6 was released by Qwen on Apr 16 2026.

Qwen3.6 leads on SWE-Bench Verified — Claude 3.5 Sonnet 33.4% vs Qwen3.6 73.4%.

Qwen3.6 leads on GPQA Diamond — Claude 3.5 Sonnet 59.4% vs Qwen3.6 86%.

Claude 3.5 Sonnet is a proprietary model released by Anthropic. Qwen3.6 is an open weight model released by Qwen.

Other comparisons