Claude Sonnet 4.5vsQwen3.6

Claude Sonnet 4.5
Qwen3.6
Specifications
Parameters
35B
Context window
256k
Benchmarks
Nonsense detection
BullshitBench v2
79%
Agentic coding
SWE-Bench Pro
49.5%
Coding
SWE-Bench Verified
77.2%Best
73.4%
Multilingual coding
SWE-Bench Multilingual
67.2%
Next.js coding
Next.js Evals
50%
Agentic terminal coding
Terminal-Bench 2.0
51.5%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
21.4%
Science
GPQA Diamond
83.4%
86%Best
Chart reasoning
CharXiv Reasoning
78%
Multimodal reasoning
MMMU-Pro
75.3%
Multimodal
MMMU
68%
81.7%Best
Overview
CompanyAnthropicQwen
Release dateSep 29 2025Apr 16 2026
AccessProprietaryOpen Weight

Which is better: Claude Sonnet 4.5 or Qwen3.6?

Qwen3.6 leads Claude Sonnet 4.5 on 2 of the 3 benchmarks they both report (SWE-Bench Verified, GPQA Diamond, MMMU). Claude Sonnet 4.5 shipped 199 days before Qwen3.6, so benchmark comparisons should account for the intervening progress.

Claude Sonnet 4.5 is proprietary, while Qwen3.6 is open weight.

On SWE-Bench Verified, Claude Sonnet 4.5 leads at 77.2% vs Qwen3.6 at 73.4%. On GPQA Diamond, Qwen3.6 leads at 86% vs Claude Sonnet 4.5 at 83.4%. On MMMU, Qwen3.6 leads at 81.7% vs Claude Sonnet 4.5 at 68%.

Frequently asked questions

Claude Sonnet 4.5 was released by Anthropic on Sep 29 2025.

Qwen3.6 was released by Qwen on Apr 16 2026.

Claude Sonnet 4.5 leads on SWE-Bench Verified — Claude Sonnet 4.5 77.2% vs Qwen3.6 73.4%.

Qwen3.6 leads on GPQA Diamond — Claude Sonnet 4.5 83.4% vs Qwen3.6 86%.

Claude Sonnet 4.5 is a proprietary model released by Anthropic. Qwen3.6 is an open weight model released by Qwen.

Other comparisons