Compare AI models

Specifications
Context window
1M1.05M
API pricing
Input price
$5.00$2.00
Output price
$25.00$10.00
Cached input price
$0.50$0.10
Cheapest input
$5.00Amazon Bedrock$1.00OpenAI
Cheapest output
$25.00Amazon Bedrock$5.00OpenAI
Benchmarks
BullshitBench v2
73%65%
DeepSWE 1.1
68.8%75.2%
Terminal-Bench-Science 0.1
29%57.1%
OSWorld 2.0
74%71.4%
AutomationBench
26.9%36.2%
Overview
CompanyAnthropicOpenAI
Release dateJul 24 2026Sep 29 2026
AccessClosedClosed
Model detailsView modelView model

Frequently asked questions

GPT-6.1 Sol leads Claude Opus 5 on 3 of the 5 benchmarks they both report. GPT-6.1 Sol is cheaper on both input and output: $2.00 vs $5.00 per million input tokens, and $10.00 vs $25.00 per million output tokens. Claude Opus 5 shipped 67 days before GPT-6.1 Sol, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Opus 5) vs 1.05M (GPT-6.1 Sol).

On BullshitBench v2, Claude Opus 5 leads at 73% vs GPT-6.1 Sol at 65%. On DeepSWE 1.1, GPT-6.1 Sol leads at 75.2% vs Claude Opus 5 at 68.8%. On Terminal-Bench-Science 0.1, GPT-6.1 Sol leads at 57.1% vs Claude Opus 5 at 29%. On OSWorld 2.0, Claude Opus 5 leads at 74% vs GPT-6.1 Sol at 71.4%. On AutomationBench, GPT-6.1 Sol leads at 36.2% vs Claude Opus 5 at 26.9%.