Claude Opus 4.1vsQwen3.6

Claude Opus 4.1
Qwen3.6
Specifications
Parameters
35B
Context window
256k
Benchmarks
Nonsense detection
BullshitBench v2
43%
Agentic coding
SWE-Bench Pro
49.5%
Coding
SWE-Bench Verified
74.5%Best
73.4%
Multilingual coding
SWE-Bench Multilingual
67.2%
Agentic terminal coding
Terminal-Bench 2.0
51.5%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
21.4%
Science
GPQA Diamond
80.9%
86%Best
Chart reasoning
CharXiv Reasoning
78%
Multimodal reasoning
MMMU-Pro
75.3%
Multimodal
MMMU
81.7%
Overview
CompanyAnthropicQwen
Release dateAug 5 2025Apr 16 2026
AccessProprietaryOpen Weight

Which is better: Claude Opus 4.1 or Qwen3.6?

Claude Opus 4.1 and Qwen3.6 are evenly matched across the 2 benchmarks they both report (SWE-Bench Verified, GPQA Diamond). Claude Opus 4.1 shipped 254 days before Qwen3.6, so benchmark comparisons should account for the intervening progress.

Claude Opus 4.1 is proprietary, while Qwen3.6 is open weight.

On SWE-Bench Verified, Claude Opus 4.1 leads at 74.5% vs Qwen3.6 at 73.4%. On GPQA Diamond, Qwen3.6 leads at 86% vs Claude Opus 4.1 at 80.9%.

Frequently asked questions

Claude Opus 4.1 was released by Anthropic on Aug 5 2025.

Qwen3.6 was released by Qwen on Apr 16 2026.

Claude Opus 4.1 leads on SWE-Bench Verified — Claude Opus 4.1 74.5% vs Qwen3.6 73.4%.

Qwen3.6 leads on GPQA Diamond — Claude Opus 4.1 80.9% vs Qwen3.6 86%.

Claude Opus 4.1 is a proprietary model released by Anthropic. Qwen3.6 is an open weight model released by Qwen.

Other comparisons