Compare AI models

Specifications
Context window
1M—
API pricing
Input price
$5.00$2.00
Output price
$25.00$10.00
Cached input price
$0.50$0.10
Cheapest input
$5.00Amazon Bedrock—
Cheapest output
$25.00Amazon Bedrock—
Benchmarks
Gray Swan IPI · k = 15
5.5%0.7%
Terminal-Bench 4.0
23.64%57.4%
Finance Agent v2
53.9%65.4%
Harvey's Legal Agent Benchmark
9.58%19.6%
Overview
CompanyAnthropicGoogle
Release dateMay 28 2026Sep 30 2026
AccessClosedClosed
Model detailsView modelView model

Frequently asked questions

Gemini 4 Argon leads Claude Opus 4.8 on 4 of the 4 benchmarks they both report (Gray Swan IPI, Terminal-Bench 4.0, Finance Agent v2, Harvey's Legal Agent Benchmark). Gemini 4 Argon is cheaper on both input and output: $2.00 vs $5.00 per million input tokens, and $10.00 vs $25.00 per million output tokens. Claude Opus 4.8 shipped 125 days before Gemini 4 Argon, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On Gray Swan IPI · k = 15, Gemini 4 Argon leads at 0.7% vs Claude Opus 4.8 at 5.5%. On Terminal-Bench 4.0, Gemini 4 Argon leads at 57.4% vs Claude Opus 4.8 at 23.64%. On Finance Agent v2, Gemini 4 Argon leads at 65.4% vs Claude Opus 4.8 at 53.9%. On Harvey's Legal Agent Benchmark, Gemini 4 Argon leads at 19.6% vs Claude Opus 4.8 at 9.58%.