Compare AI models

Specifications
Context window
—1M
API pricing
Input price
$2.00$0.75
Output price
$10.00$3.75
Cached input price
$0.20$0.075
Cheapest input
$2.00Amazon Bedrock$0.375Google
Cheapest output
$10.00Amazon Bedrock$1.875Google
Benchmarks
BullshitBench v2
80%35%
Terminal-Bench 4.0
12.42%11.21%
Terminal-Bench 2.1
80.4%85.8%
threejseval
12461480
Overview
CompanyAnthropicGoogle
Release dateJun 30 2026Aug 13 2026
AccessClosedClosed
Model detailsView modelView model

Frequently asked questions

Claude Sonnet 5 and Gemini 3.7 Flash are evenly matched across the 4 benchmarks they both report (BullshitBench v2, Terminal-Bench 4.0, Terminal-Bench 2.1, threejseval). Gemini 3.7 Flash is cheaper on both input and output: $0.75 vs $2.00 per million input tokens, and $3.75 vs $10.00 per million output tokens. Claude Sonnet 5 shipped 44 days before Gemini 3.7 Flash, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Claude Sonnet 5 leads at 80% vs Gemini 3.7 Flash at 35%. On Terminal-Bench 4.0, Claude Sonnet 5 leads at 12.42% vs Gemini 3.7 Flash at 11.21%. On Terminal-Bench 2.1, Gemini 3.7 Flash leads at 85.8% vs Claude Sonnet 5 at 80.4%. On threejseval, Gemini 3.7 Flash leads at 1480 vs Claude Sonnet 5 at 1246.