Compare AI models

Specifications
Context window
1M1M
API pricing
Input price
$5.00$0.75
Output price
$25.00$3.75
Cached input price
$0.50$0.075
Cheapest input
$5.00Amazon Bedrock$0.375Google
Cheapest output
$25.00Amazon Bedrock$1.875Google
Benchmarks
BullshitBench v2
83%35%
ProgramBench
0%0%
Terminal-Bench 2.1
66.1%85.8%
ARC-AGI-2
75.8%84.6%
CharXiv Reasoning
82.1%84.5%
MRCR v2 (8-needle) · 128k average
59.3%97%
Overview
CompanyAnthropicGoogle
Release dateApr 16 2026Aug 13 2026
AccessClosedClosed
Model detailsView modelView model

Frequently asked questions

Gemini 3.7 Flash leads Claude Opus 4.7 on 4 of the 6 benchmarks they both report. Gemini 3.7 Flash is cheaper on both input and output: $0.75 vs $5.00 per million input tokens, and $3.75 vs $25.00 per million output tokens. Claude Opus 4.7 shipped 119 days before Gemini 3.7 Flash, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Opus 4.7) vs 1M (Gemini 3.7 Flash).

On BullshitBench v2, Claude Opus 4.7 leads at 83% vs Gemini 3.7 Flash at 35%. On ProgramBench, both models score 0%. On Terminal-Bench 2.1, Gemini 3.7 Flash leads at 85.8% vs Claude Opus 4.7 at 66.1%. On ARC-AGI-2, Gemini 3.7 Flash leads at 84.6% vs Claude Opus 4.7 at 75.8%. On CharXiv Reasoning, Gemini 3.7 Flash leads at 84.5% vs Claude Opus 4.7 at 82.1%. On MRCR v2 (8-needle) · 128k average, Gemini 3.7 Flash leads at 97% vs Claude Opus 4.7 at 59.3%.