Claude Opus 4.1vsQwen3.5

Claude Opus 4.1
Qwen3.5
Specifications
Parameters
397B
Context window
1M
Benchmarks
Nonsense detection
BullshitBench v2
43%
Coding
SWE-Bench Verified
74.5%
76.4%Best
Multilingual coding
SWE-Bench Multilingual
69.3%
Agentic terminal coding
Terminal-Bench 2.0
52.5%
Web browsing
BrowseComp
69%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
28.7%
Science
GPQA Diamond
80.9%
88.4%Best
Agentic computer use
OSWorld-Verified
62.2%
Chart reasoning
CharXiv Reasoning
80.8%
Multimodal reasoning
MMMU-Pro
79%
Multimodal
MMMU
85%
Overview
CompanyAnthropicQwen
Release dateAug 5 2025Feb 16 2026
AccessProprietaryOpen Weight

Which is better: Claude Opus 4.1 or Qwen3.5?

Qwen3.5 leads Claude Opus 4.1 on 2 of the 2 benchmarks they both report (SWE-Bench Verified, GPQA Diamond). Claude Opus 4.1 shipped 195 days before Qwen3.5, so benchmark comparisons should account for the intervening progress.

Claude Opus 4.1 is proprietary, while Qwen3.5 is open weight.

On SWE-Bench Verified, Qwen3.5 leads at 76.4% vs Claude Opus 4.1 at 74.5%. On GPQA Diamond, Qwen3.5 leads at 88.4% vs Claude Opus 4.1 at 80.9%.

Frequently asked questions

Claude Opus 4.1 was released by Anthropic on Aug 5 2025.

Qwen3.5 was released by Qwen on Feb 16 2026.

Qwen3.5 leads on SWE-Bench Verified — Claude Opus 4.1 74.5% vs Qwen3.5 76.4%.

Qwen3.5 leads on GPQA Diamond — Claude Opus 4.1 80.9% vs Qwen3.5 88.4%.

Claude Opus 4.1 is a proprietary model released by Anthropic. Qwen3.5 is an open weight model released by Qwen.

Other comparisons