Claude 3.5 HaikuvsClaude Opus 5

Claude 3.5 Haiku
Claude Opus 5
Specifications
Context window
1M
Benchmarks
Nonsense detection
BullshitBench v2
50%
73%Best
Prompt injection robustness
Gray Swan IPI · k = 1
0.2%
Prompt injection robustness
Gray Swan IPI · k = 10
1.6%
Prompt injection robustness
Gray Swan IPI · k = 15
2%
Coding
SWE-Bench Verified
40.6%
Agentic coding
CursorBench v3.2
70%
Agentic coding
DeepSWE 1.1
68.8%
Agentic coding
FrontierCode v1.1 (Main) · main split
53.4%
Next.js coding
Next.js Evals
88%
Supabase coding
Supabase Evals · with skills
95.5%
Supabase coding
Supabase Evals · no skills
90.9%
Agentic computer work
Frontier-Bench v0.1
43.3%
Web browsing
BrowseComp
90.8%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
56.3%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
64.7%
Novel problem-solving
ARC-AGI-3
30.2%
Biology
BioMysteryBench · hard
49.4%
Biology
BioMysteryBench · human solved
90.1%
Science
GPQA Diamond
41.6%
Agentic computer use
OSWorld 2.0
70.6%
Business workflows
AutomationBench
26%
Agentic legal work
Harvey's Legal Agent Benchmark (Held-out)
11.7%
Health
HealthBench Professional
59.8%
Knowledge work
GDPval-AA v2
1861
Community preference
Arena Elo (Text)
1495
Community preference (code)
Arena Elo (Code)
1663
Overview
CompanyAnthropicAnthropic
Release dateOct 22 2024Jul 24 2026
AccessProprietaryProprietary

Which is better: Claude 3.5 Haiku or Claude Opus 5?

Claude Opus 5 leads Claude 3.5 Haiku on 1 of the 1 benchmark they both report (BullshitBench v2). Claude 3.5 Haiku shipped 640 days before Claude Opus 5, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Claude Opus 5 leads at 73% vs Claude 3.5 Haiku at 50%.

Frequently asked questions

Claude 3.5 Haiku was released by Anthropic on Oct 22 2024.

Claude Opus 5 was released by Anthropic on Jul 24 2026.

Other comparisons