Compare AI models

Specifications
Context window
1M—
API pricing
Input price
$4.00$2.00
Output price
$20.00$10.00
Cached input price
$0.20$0.10
Cheapest input
$4.00Amazon Bedrock—
Cheapest output
$20.00Amazon Bedrock—
Benchmarks
Terminal-Bench 4.0
66.4%57.4%
Terminal-Bench-Science 0.1
58.7%57.6%
OSWorld 2.0
81.8%69.2%
AutomationBench
40%51.3%
Overview
CompanyAnthropicGoogle
Release dateSep 22 2026Sep 30 2026
AccessClosedClosed
Model detailsView modelView model

Frequently asked questions

Claude Opus 5.5 leads Gemini 4 Argon on 3 of the 4 benchmarks they both report (Terminal-Bench 4.0, Terminal-Bench-Science 0.1, OSWorld 2.0, AutomationBench). Gemini 4 Argon is cheaper on both input and output: $2.00 vs $4.00 per million input tokens, and $10.00 vs $20.00 per million output tokens. Claude Opus 5.5 shipped 8 days before Gemini 4 Argon, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On Terminal-Bench 4.0, Claude Opus 5.5 leads at 66.4% vs Gemini 4 Argon at 57.4%. On Terminal-Bench-Science 0.1, Claude Opus 5.5 leads at 58.7% vs Gemini 4 Argon at 57.6%. On OSWorld 2.0, Claude Opus 5.5 leads at 81.8% vs Gemini 4 Argon at 69.2%. On AutomationBench, Gemini 4 Argon leads at 51.3% vs Claude Opus 5.5 at 40%.