Compare AI models

Specifications
Context window
1M—
API pricing
Input price
$5.00$2.00
Output price
$25.00$10.00
Cached input price
$0.50$0.10
Cheapest input
$5.00Amazon Bedrock—
Cheapest output
$25.00Amazon Bedrock—
Benchmarks
Gray Swan IPI · k = 15
2%0.7%
DeepSWE 1.1
68.8%77.9%
Terminal-Bench 4.0
51.82%57.4%
Terminal-Bench-Science 0.1
29%57.6%
OSWorld 2.0
74%69.2%
AutomationBench
26.9%51.3%
Overview
CompanyAnthropicGoogle
Release dateJul 24 2026Sep 30 2026
AccessClosedClosed
Model detailsView modelView model

Frequently asked questions

Gemini 4 Argon leads Claude Opus 5 on 5 of the 6 benchmarks they both report. Gemini 4 Argon is cheaper on both input and output: $2.00 vs $5.00 per million input tokens, and $10.00 vs $25.00 per million output tokens. Claude Opus 5 shipped 68 days before Gemini 4 Argon, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On Gray Swan IPI · k = 15, Gemini 4 Argon leads at 0.7% vs Claude Opus 5 at 2%. On DeepSWE 1.1, Gemini 4 Argon leads at 77.9% vs Claude Opus 5 at 68.8%. On Terminal-Bench 4.0, Gemini 4 Argon leads at 57.4% vs Claude Opus 5 at 51.82%. On Terminal-Bench-Science 0.1, Gemini 4 Argon leads at 57.6% vs Claude Opus 5 at 29%. On OSWorld 2.0, Claude Opus 5 leads at 74% vs Gemini 4 Argon at 69.2%. On AutomationBench, Gemini 4 Argon leads at 51.3% vs Claude Opus 5 at 26.9%.