Claude Opus 5vsClaude Opus 5.5

Claude Opus 5
Claude Opus 5.5
Specifications
Context window
1M1M
API pricing
Input price
$5.00$4.00
Output price
$25.00$20.00
Cached input price
$0.50$0.20
Cheapest input
$5.00Amazon Bedrock
Cheapest output
$25.00Amazon Bedrock
Benchmarks
FrontierCode v1.1 (Main) · main split
48%54.4%
Terminal-Bench 4.0
51.82%66.4%
Terminal-Bench-Science 0.1
29%58.7%
Humanity's Last Exam · with tools
63.6%67.7%
OSWorld 2.0
74%81.8%
AutomationBench
26.9%40%
GDPval-AA v2.1
17081846
Chartography · with tools
83.4%89%
Benchmarks
BullshitBench v2
73%
Gray Swan IPI · k = 1
0.2%
Gray Swan IPI · k = 10
1.6%
Gray Swan IPI · k = 15
2%
ProgramBench
4.5%
SWE-Bench Pro
79.2%
SWE-Bench Verified
96%
SWE-Bench Multilingual
89.5%
SWE-Bench Multimodal
59.4%
DeepSWE 1.1
68.8%
Next.js Evals
94%
Supabase Evals · with skills
92.8%
Supabase Evals · no skills
89.9%
Frontier-Bench v0.1
43.3%
BrowseComp
90.8%
Humanity's Last Exam · no tools
56.3%
ARC-AGI-3
30.2%
ARC-AGI-2
90.4%
BioMysteryBench · hard
49.4%
BioMysteryBench · human solved
90.1%
Harvey's Legal Agent Benchmark (Held-out)
11.7%
HealthBench Professional
59.8%
GDPval-AA v2
1861
AA-Briefcase
1685
threejseval
1776
Overview
CompanyAnthropicAnthropic
Release dateJul 24 2026Sep 22 2026
AccessProprietaryProprietary

Other comparisons

Claude Opus 5vsGPT-6 AstraClaude Opus 5.5vsGPT-6 AstraClaude Opus 5vsGemini 3.8 FlashClaude Opus 5.5vsGemini 3.8 FlashClaude Opus 5vsMuse Spark 1.3Claude Opus 5.5vsMuse Spark 1.3Claude Opus 5vsGrok 4.7Claude Opus 5.5vsGrok 4.7Claude Opus 5vsDeepSeek-V4.1-FlashClaude Opus 5.5vsDeepSeek-V4.1-FlashClaude Opus 5vsMistral Medium 3.5Claude Opus 5.5vsMistral Medium 3.5

Frequently asked questions

Claude Opus 5.5 leads Claude Opus 5 on 8 of the 8 benchmarks they both report. Claude Opus 5.5 is cheaper on both input and output: $4.00 vs $5.00 per million input tokens, and $20.00 vs $25.00 per million output tokens. Claude Opus 5 shipped 60 days before Claude Opus 5.5, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Claude Opus 5) vs 1M (Claude Opus 5.5).

On FrontierCode v1.1 (Main) · main split, Claude Opus 5.5 leads at 54.4% vs Claude Opus 5 at 48%. On Terminal-Bench 4.0, Claude Opus 5.5 leads at 66.4% vs Claude Opus 5 at 51.82%. On Terminal-Bench-Science 0.1, Claude Opus 5.5 leads at 58.7% vs Claude Opus 5 at 29%. On Humanity's Last Exam · with tools, Claude Opus 5.5 leads at 67.7% vs Claude Opus 5 at 63.6%. On OSWorld 2.0, Claude Opus 5.5 leads at 81.8% vs Claude Opus 5 at 74%. On AutomationBench, Claude Opus 5.5 leads at 40% vs Claude Opus 5 at 26.9%. On GDPval-AA v2.1, Claude Opus 5.5 leads at 1846 vs Claude Opus 5 at 1708. On Chartography · with tools, Claude Opus 5.5 leads at 89% vs Claude Opus 5 at 83.4%.