Compare AI models

API pricing
Input price
$3.00$1.50
Output price
$15.00$9.00
Cached input price
$0.30$0.15
Cheapest input
$3.00Anthropic$0.75Google
Cheapest output
$15.00Anthropic$4.50Google
Benchmarks
BullshitBench v2
91%20%
ProgramBench
0%0%
DeepSWE 1.1
30%37%
MCP Atlas
69.5%83.6%
BU Bench
62%58%
Humanity's Last Exam · no tools
33.2%40.2%
Humanity's Last Exam · with tools
49%40.2%
ARC-AGI-2
58.3%72.1%
OSWorld-Verified
72.5%78.4%
Finance Agent v2
51%57.9%
GDPval-AA
16761656
CharXiv Reasoning
72.4%84.2%
MMMU-Pro
74.5%83.6%
Blueprint-Bench 2
6.7%33.6%
MRCR v2 (8-needle) · 128k average
84.9%77.3%
Overview
CompanyAnthropicGoogle
Release dateFeb 17 2026May 19 2026
AccessClosedClosed
Model detailsView modelView model

Frequently asked questions

Gemini 3.5 Flash leads Claude Sonnet 4.6 on 9 of the 15 benchmarks they both report. Gemini 3.5 Flash is cheaper on both input and output: $1.50 vs $3.00 per million input tokens, and $9.00 vs $15.00 per million output tokens. Claude Sonnet 4.6 shipped 91 days before Gemini 3.5 Flash, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Claude Sonnet 4.6 leads at 91% vs Gemini 3.5 Flash at 20%. On ProgramBench, both models score 0%. On DeepSWE 1.1, Gemini 3.5 Flash leads at 37% vs Claude Sonnet 4.6 at 30%. On MCP Atlas, Gemini 3.5 Flash leads at 83.6% vs Claude Sonnet 4.6 at 69.5%. On BU Bench, Claude Sonnet 4.6 leads at 62% vs Gemini 3.5 Flash at 58%. On Humanity's Last Exam · no tools, Gemini 3.5 Flash leads at 40.2% vs Claude Sonnet 4.6 at 33.2%. On Humanity's Last Exam · with tools, Claude Sonnet 4.6 leads at 49% vs Gemini 3.5 Flash at 40.2%. On ARC-AGI-2, Gemini 3.5 Flash leads at 72.1% vs Claude Sonnet 4.6 at 58.3%. On OSWorld-Verified, Gemini 3.5 Flash leads at 78.4% vs Claude Sonnet 4.6 at 72.5%. On Finance Agent v2, Gemini 3.5 Flash leads at 57.9% vs Claude Sonnet 4.6 at 51%. On GDPval-AA, Claude Sonnet 4.6 leads at 1676 vs Gemini 3.5 Flash at 1656. On CharXiv Reasoning, Gemini 3.5 Flash leads at 84.2% vs Claude Sonnet 4.6 at 72.4%. On MMMU-Pro, Gemini 3.5 Flash leads at 83.6% vs Claude Sonnet 4.6 at 74.5%. On Blueprint-Bench 2, Gemini 3.5 Flash leads at 33.6% vs Claude Sonnet 4.6 at 6.7%. On MRCR v2 (8-needle) · 128k average, Claude Sonnet 4.6 leads at 84.9% vs Gemini 3.5 Flash at 77.3%.