Compare AI models

API pricing
Cheapest input
—$0.05DeepInfra
Cheapest output
—$0.10DeepInfra
Benchmarks
BullshitBench v2
46%3%
Overview
CompanyAnthropicGoogle
Release dateJun 20 2024Mar 12 2025
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

Claude 3.5 Sonnet leads Gemma 3 on 1 of the 1 benchmark they both report (BullshitBench v2). Claude 3.5 Sonnet shipped 265 days before Gemma 3, so benchmark comparisons should account for the intervening progress.

Claude 3.5 Sonnet is closed, while Gemma 3 is open weight.

On BullshitBench v2, Claude 3.5 Sonnet leads at 46% vs Gemma 3 at 3%.