Claude 3.5 HaikuvsQwen3.5

Claude 3.5 Haiku
Qwen3.5
Specifications
Parameters
397B
Context window
1M
Benchmarks
Nonsense detection
BullshitBench v2
50%
Coding
SWE-Bench Verified
40.6%
76.4%Best
Multilingual coding
SWE-Bench Multilingual
69.3%
Agentic terminal coding
Terminal-Bench 2.0
52.5%
Web browsing
BrowseComp
69%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
28.7%
Science
GPQA Diamond
41.6%
88.4%Best
Agentic computer use
OSWorld-Verified
62.2%
Chart reasoning
CharXiv Reasoning
80.8%
Multimodal reasoning
MMMU-Pro
79%
Multimodal
MMMU
85%
Overview
CompanyAnthropicQwen
Release dateOct 22 2024Feb 16 2026
AccessProprietaryOpen Weight

Which is better: Claude 3.5 Haiku or Qwen3.5?

Qwen3.5 leads Claude 3.5 Haiku on 2 of the 2 benchmarks they both report (SWE-Bench Verified, GPQA Diamond). Claude 3.5 Haiku shipped 482 days before Qwen3.5, so benchmark comparisons should account for the intervening progress.

Claude 3.5 Haiku is proprietary, while Qwen3.5 is open weight.

On SWE-Bench Verified, Qwen3.5 leads at 76.4% vs Claude 3.5 Haiku at 40.6%. On GPQA Diamond, Qwen3.5 leads at 88.4% vs Claude 3.5 Haiku at 41.6%.

Frequently asked questions

Claude 3.5 Haiku was released by Anthropic on Oct 22 2024.

Qwen3.5 was released by Qwen on Feb 16 2026.

Qwen3.5 leads on SWE-Bench Verified — Claude 3.5 Haiku 40.6% vs Qwen3.5 76.4%.

Qwen3.5 leads on GPQA Diamond — Claude 3.5 Haiku 41.6% vs Qwen3.5 88.4%.

Claude 3.5 Haiku is a proprietary model released by Anthropic. Qwen3.5 is an open weight model released by Qwen.

Other comparisons