Claude Sonnet 4.6vsQwen3.6

Claude Sonnet 4.6
Qwen3.6
Specifications
Parameters
35B
Context window
256k
Benchmarks
Nonsense detection
BullshitBench v2
91%
Agentic coding
SWE-Bench Pro
49.5%
Coding
SWE-Bench Verified
79.6%Best
73.4%
Multilingual coding
SWE-Bench Multilingual
67.2%
Agentic coding
CursorBench v3.1
49%
Agentic coding
DeepSWE 1.1
30%
Next.js coding
Next.js Evals
58%
Agentic terminal coding
Terminal-Bench 2.0
51.5%
Multi-step tool use
MCP Atlas
69.5%
Browser agent
BU Bench
62%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
33.2%Best
21.4%
Abstract reasoning
ARC-AGI-2
58.3%
Science
GPQA Diamond
89.9%Best
86%
Agentic computer use
OSWorld-Verified
72.5%
Agentic financial analysis
Finance Agent v2
51%
Knowledge work
GDPval-AA
1676
Chart reasoning
CharXiv Reasoning
72.4%
78%Best
Multimodal reasoning
MMMU-Pro
74.5%
75.3%Best
Multimodal
MMMU
81.7%
Spatial reasoning
Blueprint-Bench 2
6.7%
Long context
MRCR v2 (8-needle) · 128k average
84.9%
Community preference (code)
Arena Elo (Code)
1521
Overview
CompanyAnthropicQwen
Release dateFeb 17 2026Apr 16 2026
AccessProprietaryOpen Weight

Which is better: Claude Sonnet 4.6 or Qwen3.6?

Claude Sonnet 4.6 leads Qwen3.6 on 3 of the 5 benchmarks they both report. Claude Sonnet 4.6 shipped 58 days before Qwen3.6, so benchmark comparisons should account for the intervening progress.

Claude Sonnet 4.6 is proprietary, while Qwen3.6 is open weight.

On SWE-Bench Verified, Claude Sonnet 4.6 leads at 79.6% vs Qwen3.6 at 73.4%. On Humanity's Last Exam · no tools, Claude Sonnet 4.6 leads at 33.2% vs Qwen3.6 at 21.4%. On GPQA Diamond, Claude Sonnet 4.6 leads at 89.9% vs Qwen3.6 at 86%. On CharXiv Reasoning, Qwen3.6 leads at 78% vs Claude Sonnet 4.6 at 72.4%. On MMMU-Pro, Qwen3.6 leads at 75.3% vs Claude Sonnet 4.6 at 74.5%.

Frequently asked questions

Claude Sonnet 4.6 was released by Anthropic on Feb 17 2026.

Qwen3.6 was released by Qwen on Apr 16 2026.

Claude Sonnet 4.6 leads on SWE-Bench Verified — Claude Sonnet 4.6 79.6% vs Qwen3.6 73.4%.

Claude Sonnet 4.6 leads on Humanity's Last Exam · no tools — Claude Sonnet 4.6 33.2% vs Qwen3.6 21.4%.

Claude Sonnet 4.6 is a proprietary model released by Anthropic. Qwen3.6 is an open weight model released by Qwen.

Other comparisons