Compare AI models

Specifications
Context window
1M1.05M
API pricing
Input price
$4.00$2.00
Output price
$20.00$10.00
Cached input price
$0.20$0.10
Cheapest input
$4.00Amazon Bedrock$1.00OpenAI
Cheapest output
$20.00Amazon Bedrock$5.00OpenAI
Benchmarks
BullshitBench v2
63%65%
Terminal-Bench-Science 0.1
58.7%57.1%
OSWorld 2.0
81.8%71.4%
AutomationBench
40%36.2%
Overview
CompanyAnthropicOpenAI
Release dateSep 22 2026Sep 29 2026
AccessClosedClosed
Model detailsView modelView model

Frequently asked questions

Claude Opus 5.5 leads GPT-6.1 Sol on 3 of the 4 benchmarks they both report (BullshitBench v2, Terminal-Bench-Science 0.1, OSWorld 2.0, AutomationBench). GPT-6.1 Sol is cheaper on both input and output: $2.00 vs $4.00 per million input tokens, and $10.00 vs $20.00 per million output tokens. Claude Opus 5.5 shipped 7 days before GPT-6.1 Sol, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Opus 5.5) vs 1.05M (GPT-6.1 Sol).

On BullshitBench v2, GPT-6.1 Sol leads at 65% vs Claude Opus 5.5 at 63%. On Terminal-Bench-Science 0.1, Claude Opus 5.5 leads at 58.7% vs GPT-6.1 Sol at 57.1%. On OSWorld 2.0, Claude Opus 5.5 leads at 81.8% vs GPT-6.1 Sol at 71.4%. On AutomationBench, Claude Opus 5.5 leads at 40% vs GPT-6.1 Sol at 36.2%.