Compare AI models

Specifications
Context window
1M—
API pricing
Input price
$0.75$1.25
Output price
$3.75$2.50
Cached input price
$0.075$0.20
Cheapest input
$0.375Google—
Cheapest output
$1.875Google—
Benchmarks
BullshitBench v2
35%56%
ARC-AGI-2
84.6%53.3%
Overview
CompanyGoogleSpaceXAI
Release dateAug 13 2026Feb 17 2026
AccessClosedClosed
Model detailsView modelView model

Frequently asked questions

Gemini 3.7 Flash and Grok 4.20 Beta are evenly matched across the 2 benchmarks they both report (BullshitBench v2, ARC-AGI-2). Gemini 3.7 Flash is cheaper on input: $0.75 vs $1.25 per million tokens. Grok 4.20 Beta is cheaper on output: $2.50 vs $3.75 per million tokens. Figures are base-tier rates. Grok 4.20 Beta shipped 177 days before Gemini 3.7 Flash, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Grok 4.20 Beta leads at 56% vs Gemini 3.7 Flash at 35%. On ARC-AGI-2, Gemini 3.7 Flash leads at 84.6% vs Grok 4.20 Beta at 53.3%.