Claude Opus 4.7vsQwen3.5

Claude Opus 4.7
Qwen3.5
Specifications
Parameters
397B
Context window
1M
1M
Benchmarks
Nonsense detection
BullshitBench v2
83%
Agentic coding
SWE-Bench Pro
64.3%
Coding
SWE-Bench Verified
87.6%Best
76.4%
Multilingual coding
SWE-Bench Multilingual
80.5%Best
69.3%
Agentic coding
CursorBench v3.1
64.8%
Next.js coding
Next.js Evals
75%
Agentic terminal coding
Terminal-Bench 2.1
66.1%
Agentic terminal coding
Terminal-Bench 2.0
69.4%Best
52.5%
Multi-step tool use
MCP Atlas
79.1%
Web browsing
BrowseComp
79.3%Best
69%
Cybersecurity
CyberGym
73.1%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
46.9%Best
28.7%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
54.7%
Abstract reasoning
ARC-AGI-2
75.8%
Advanced math
FrontierMath · Tier 1–3
43.8%
Advanced math
FrontierMath · Tier 4
22.9%
Science
GPQA Diamond
94.2%Best
88.4%
Agentic computer use
OSWorld-Verified
78%Best
62.2%
Agentic financial analysis
Finance Agent v2
51.5%
Knowledge work
GDPval-AA
1753
Knowledge work
GDPval (win/tie rate)
80.3%
Chart reasoning
CharXiv Reasoning
82.1%Best
80.8%
Multimodal reasoning
MMMU-Pro
75.2%
79%Best
Multimodal
MMMU
85%
Spatial reasoning
Blueprint-Bench 2
24.5%
Long context
MRCR v2 (8-needle) · 128k average
59.3%
Community preference
Arena Elo (Text)
1503
Community preference (code)
Arena Elo (Code)
1557
Overview
CompanyAnthropicQwen
Release dateApr 16 2026Feb 16 2026
AccessProprietaryOpen Weight

Which is better: Claude Opus 4.7 or Qwen3.5?

Claude Opus 4.7 leads Qwen3.5 on 8 of the 9 benchmarks they both report. Qwen3.5 shipped 59 days before Claude Opus 4.7, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Opus 4.7) vs 1M (Qwen3.5). Claude Opus 4.7 is proprietary, while Qwen3.5 is open weight.

On SWE-Bench Verified, Claude Opus 4.7 leads at 87.6% vs Qwen3.5 at 76.4%. On SWE-Bench Multilingual, Claude Opus 4.7 leads at 80.5% vs Qwen3.5 at 69.3%. On Terminal-Bench 2.0, Claude Opus 4.7 leads at 69.4% vs Qwen3.5 at 52.5%. On BrowseComp, Claude Opus 4.7 leads at 79.3% vs Qwen3.5 at 69%. On Humanity's Last Exam · no tools, Claude Opus 4.7 leads at 46.9% vs Qwen3.5 at 28.7%. On GPQA Diamond, Claude Opus 4.7 leads at 94.2% vs Qwen3.5 at 88.4%. On OSWorld-Verified, Claude Opus 4.7 leads at 78% vs Qwen3.5 at 62.2%. On CharXiv Reasoning, Claude Opus 4.7 leads at 82.1% vs Qwen3.5 at 80.8%. On MMMU-Pro, Qwen3.5 leads at 79% vs Claude Opus 4.7 at 75.2%.

Frequently asked questions

Claude Opus 4.7 was released by Anthropic on Apr 16 2026.

Qwen3.5 was released by Qwen on Feb 16 2026.

Claude Opus 4.7 leads on SWE-Bench Verified — Claude Opus 4.7 87.6% vs Qwen3.5 76.4%.

Claude Opus 4.7 leads on Humanity's Last Exam · no tools — Claude Opus 4.7 46.9% vs Qwen3.5 28.7%.

Claude Opus 4.7 has a 1M context window; Qwen3.5 has 1M.

Claude Opus 4.7 is a proprietary model released by Anthropic. Qwen3.5 is an open weight model released by Qwen.

Other comparisons