Compare AI models

API pricing
Cheapest input
—$0.05DeepInfra
Cheapest output
—$0.10DeepInfra
Benchmarks
BullshitBench v2
34%3%
Overview
CompanyAnthropicGoogle
Release dateMay 22 2025Mar 12 2025
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

Claude Opus 4 leads Gemma 3 on 1 of the 1 benchmark they both report (BullshitBench v2). Gemma 3 shipped 71 days before Claude Opus 4, so benchmark comparisons should account for the intervening progress.

Claude Opus 4 is closed, while Gemma 3 is open weight.

On BullshitBench v2, Claude Opus 4 leads at 34% vs Gemma 3 at 3%.