Compare AI models

Specifications
Context window
1M1.05M
API pricing
Input price
$10.00$2.00
Output price
$50.00$10.00
Cached input price
$0.25$0.10
Cheapest input
$10.00Amazon Bedrock$1.00OpenAI
Cheapest output
$50.00Amazon Bedrock$5.00OpenAI
Benchmarks
BullshitBench v2
77%65%
Terminal-Bench-Science 0.1
52.6%57.1%
OSWorld 2.0
80.7%71.4%
AutomationBench
31.4%36.2%
Overview
CompanyAnthropicOpenAI
Release dateSep 1 2026Sep 29 2026
AccessClosedClosed
Model detailsView modelView model

Frequently asked questions

Claude Fable 5.1 and GPT-6.1 Sol are evenly matched across the 4 benchmarks they both report (BullshitBench v2, Terminal-Bench-Science 0.1, OSWorld 2.0, AutomationBench). GPT-6.1 Sol is cheaper on both input and output: $2.00 vs $10.00 per million input tokens, and $10.00 vs $50.00 per million output tokens. Claude Fable 5.1 shipped 28 days before GPT-6.1 Sol, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Fable 5.1) vs 1.05M (GPT-6.1 Sol).

On BullshitBench v2, Claude Fable 5.1 leads at 77% vs GPT-6.1 Sol at 65%. On Terminal-Bench-Science 0.1, GPT-6.1 Sol leads at 57.1% vs Claude Fable 5.1 at 52.6%. On OSWorld 2.0, Claude Fable 5.1 leads at 80.7% vs GPT-6.1 Sol at 71.4%. On AutomationBench, GPT-6.1 Sol leads at 36.2% vs Claude Fable 5.1 at 31.4%.