Compare AI models

Specifications
Context window
1M—
API pricing
Input price
$10.00$2.00
Output price
$50.00$10.00
Cached input price
$0.25$0.10
Cheapest input
$10.00Amazon Bedrock—
Cheapest output
$50.00Amazon Bedrock—
Benchmarks
Terminal-Bench 4.0
57.88%57.4%
Terminal-Bench-Science 0.1
52.6%57.6%
OSWorld 2.0
80.7%69.2%
AutomationBench
31.4%51.3%
Overview
CompanyAnthropicGoogle
Release dateSep 1 2026Sep 30 2026
AccessClosedClosed
Model detailsView modelView model

Frequently asked questions

Claude Fable 5.1 and Gemini 4 Argon are evenly matched across the 4 benchmarks they both report (Terminal-Bench 4.0, Terminal-Bench-Science 0.1, OSWorld 2.0, AutomationBench). Gemini 4 Argon is cheaper on both input and output: $2.00 vs $10.00 per million input tokens, and $10.00 vs $50.00 per million output tokens. Claude Fable 5.1 shipped 29 days before Gemini 4 Argon, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On Terminal-Bench 4.0, Claude Fable 5.1 leads at 57.88% vs Gemini 4 Argon at 57.4%. On Terminal-Bench-Science 0.1, Gemini 4 Argon leads at 57.6% vs Claude Fable 5.1 at 52.6%. On OSWorld 2.0, Claude Fable 5.1 leads at 80.7% vs Gemini 4 Argon at 69.2%. On AutomationBench, Gemini 4 Argon leads at 51.3% vs Claude Fable 5.1 at 31.4%.