Compare AI models

API pricing
Input price
$1.50$1.25
Output price
$9.00$2.50
Cached input price
$0.15$0.20
Cheapest input
$0.75Google—
Cheapest output
$4.50Google—
Benchmarks
BullshitBench v2
20%56%
ARC-AGI-2
72.1%53.3%
Overview
CompanyGoogleSpaceXAI
Release dateMay 19 2026Feb 17 2026
AccessClosedClosed
Model detailsView modelView model

Frequently asked questions

Gemini 3.5 Flash and Grok 4.20 Beta are evenly matched across the 2 benchmarks they both report (BullshitBench v2, ARC-AGI-2). Grok 4.20 Beta is cheaper on both input and output: $1.25 vs $1.50 per million input tokens, and $2.50 vs $9.00 per million output tokens. Figures are base-tier rates. Grok 4.20 Beta shipped 91 days before Gemini 3.5 Flash, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Grok 4.20 Beta leads at 56% vs Gemini 3.5 Flash at 20%. On ARC-AGI-2, Gemini 3.5 Flash leads at 72.1% vs Grok 4.20 Beta at 53.3%.