Claude Opus 5vsQwen3.6

Claude Opus 5
Qwen3.6
Specifications
Parameters
35B
Context window
1M
256k
Benchmarks
Nonsense detection
BullshitBench v2
73%
Prompt injection robustness
Gray Swan IPI · k = 1
0.2%
Prompt injection robustness
Gray Swan IPI · k = 10
1.6%
Prompt injection robustness
Gray Swan IPI · k = 15
2%
Agentic coding
SWE-Bench Pro
49.5%
Coding
SWE-Bench Verified
73.4%
Multilingual coding
SWE-Bench Multilingual
67.2%
Agentic coding
CursorBench v3.2
70%
Agentic coding
DeepSWE 1.1
68.8%
Agentic coding
FrontierCode v1.1 (Main)
53.4%
Supabase coding
Supabase Evals · with skills
94.7%
Supabase coding
Supabase Evals · no skills
94.7%
Agentic computer work
Frontier-Bench v0.1
43.3%
Agentic terminal coding
Terminal-Bench 2.0
51.5%
Web browsing
BrowseComp
90.8%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
56.3%Best
21.4%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
64.7%
Novel problem-solving
ARC-AGI-3
30.2%
Biology
BioMysteryBench · hard
49.4%
Biology
BioMysteryBench · human solved
90.1%
Science
GPQA Diamond
86%
Agentic computer use
OSWorld 2.0
70.6%
Business workflows
AutomationBench
26%
Agentic legal work
Harvey's Legal Agent Benchmark (Held-out)
11.7%
Health
HealthBench Professional
59.8%
Knowledge work
GDPval-AA v2
1861
Chart reasoning
CharXiv Reasoning
78%
Multimodal reasoning
MMMU-Pro
75.3%
Multimodal
MMMU
81.7%
Community preference
Arena Elo (Text)
1495
Community preference (code)
Arena Elo (Code)
1673
Overview
CompanyAnthropicQwen
Release dateJul 24 2026Apr 16 2026
AccessProprietaryOpen Weight

Which is better: Claude Opus 5 or Qwen3.6?

Claude Opus 5 leads Qwen3.6 on 1 of the 1 benchmark they both report (Humanity's Last Exam). Qwen3.6 shipped 99 days before Claude Opus 5, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Opus 5) vs 256k (Qwen3.6). Claude Opus 5 is proprietary, while Qwen3.6 is open weight.

On Humanity's Last Exam · no tools, Claude Opus 5 leads at 56.3% vs Qwen3.6 at 21.4%.

Frequently asked questions

Claude Opus 5 was released by Anthropic on Jul 24 2026.

Qwen3.6 was released by Qwen on Apr 16 2026.

Claude Opus 5 leads on Humanity's Last Exam · no tools — Claude Opus 5 56.3% vs Qwen3.6 21.4%.

Claude Opus 5 has a 1M context window; Qwen3.6 has 256k.

Claude Opus 5 is a proprietary model released by Anthropic. Qwen3.6 is an open weight model released by Qwen.

Other comparisons