Claude 3.5 HaikuvsGPT-4o mini

Claude 3.5 Haiku
GPT-4o mini
Specifications
Context window
128k
Benchmarks
Nonsense detection
BullshitBench v2
50%Best
2%
Coding
SWE-Bench Verified
40.6%
Science
GPQA Diamond
41.6%Best
40.2%
Overview
CompanyAnthropicOpenAI
Release dateOct 22 2024Jul 18 2024
AccessProprietaryProprietary

Which is better: Claude 3.5 Haiku or GPT-4o mini?

Claude 3.5 Haiku leads GPT-4o mini on 2 of the 2 benchmarks they both report (BullshitBench v2, GPQA Diamond). GPT-4o mini shipped 96 days before Claude 3.5 Haiku, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Claude 3.5 Haiku leads at 50% vs GPT-4o mini at 2%. On GPQA Diamond, Claude 3.5 Haiku leads at 41.6% vs GPT-4o mini at 40.2%.

Frequently asked questions

Claude 3.5 Haiku was released by Anthropic on Oct 22 2024.

GPT-4o mini was released by OpenAI on Jul 18 2024.

Claude 3.5 Haiku leads on GPQA Diamond — Claude 3.5 Haiku 41.6% vs GPT-4o mini 40.2%.

Other comparisons