Claude Opus 5

Claude Opus 5 is an AI model released by Anthropic on Friday, Jul 24 2026, 24 days after Claude Sonnet 5. It has a 1M token context window. Benchmark results (shown below) cover Gray Swan IPI, FrontierCode v1.1 (Main), Frontier-Bench v0.1, Humanity's Last Exam, ARC-AGI-3, BioMysteryBench, and 13 more.

Benchmarks

Prompt injection robustness
Gray Swan IPI
0.2%
k = 1
1.6%
k = 10
2%
k = 15
k = 1
#1
k = 10
#1
k = 15
#1
Agentic coding
FrontierCode v1.1 (Main)
53.4%
Agentic computer work
Frontier-Bench v0.1
43.3%
#1
Multidisciplinary reasoning
Humanity's Last Exam
56.3%
no tools
64.7%
with tools
no tools
#1
with tools
#1
Novel problem-solving
ARC-AGI-3
30.2%
Biology
BioMysteryBench
49.4%
hard
90.1%
human solved
hard
human solved
Agentic computer use
OSWorld 2.0
70.6%
Business workflows
AutomationBench
26%
#1
Agentic legal work
Harvey's Legal Agent Benchmark (Held-out)
11.7%
Health
HealthBench Professional
59.8%
Knowledge work
GDPval-AA v2
1861
#1
Nonsense detection
BullshitBench v2
73%
#9 of 67
Agentic coding
CursorBench v3.2
70%
#2 of 14
Agentic coding
DeepSWE 1.1
68.8%
#4 of 14
Next.js coding
Next.js Evals
88%
#5 of 24
Supabase coding
Supabase Evals
94.7%
with skills
94.7%
no skills
with skills
#2 of 5
no skills
#2 of 5
Web browsing
BrowseComp
90.8%
#2 of 17
Community preference
Arena Elo (Text)
1495
#4 of 17
Community preference (code)
Arena Elo (Code)
1673
#2 of 43

About Claude Opus 5

Claude Opus 5, released July 24, 2026, brought Anthropic's Claude 5 generation to the Opus tier six weeks after Claude Fable 5 opened it — at half Fable's price, keeping Opus 4.8's $5 per million input tokens and $25 per million output. The pitch was flagship-class agentic capability at workhorse pricing: at launch it scored 43.3% on Frontier-Bench v0.1, more than double Opus 4.8's 21.1% and nearly ten points clear of Fable 5, and posted a GDPval-AA v2 Elo of 1861 for knowledge work, the best published score at the time.

The launch card leaned on breadth: 90.8% on BrowseComp for agentic search, 70.6% on OSWorld 2.0 computer use, 64.7% on Humanity's Last Exam with tools, and 30.2% on ARC-AGI-3 — which Anthropic reported as roughly three times the next best published result on the novel problem-solving benchmark. Anthropic made it the default model on Claude Max and the strongest model available on Claude Pro, positioning Opus 5 as the everyday frontier model while Fable 5 kept the edge on a handful of evaluations, including DeepSWE agentic coding and Harvey's held-out legal benchmark.

Claude Opus 5 — frequently asked questions

When was Claude Opus 5 released?
Claude Opus 5 was released by Anthropic on Friday, Jul 24 2026.
Who made Claude Opus 5?
Claude Opus 5 was built by Anthropic. AI safety company building the Claude family of models. Founded in 2021 by former OpenAI researchers.
What benchmark scores did Claude Opus 5 get?
Claude Opus 5 reports 24 tracked benchmark scores — BullshitBench v2: 73%; Gray Swan IPI (k = 1): 0.2%; Gray Swan IPI (k = 10): 1.6%; Gray Swan IPI (k = 15): 2%; CursorBench v3.2: 70%; DeepSWE 1.1: 68.8%; FrontierCode v1.1 (Main): 53.4%; Next.js Evals: 88%; Supabase Evals (with skills): 94.7%; Supabase Evals (no skills): 94.7%; Frontier-Bench v0.1: 43.3%; BrowseComp: 90.8%; Humanity's Last Exam (no tools): 56.3%; Humanity's Last Exam (with tools): 64.7%; ARC-AGI-3: 30.2%; BioMysteryBench (hard): 49.4%; BioMysteryBench (human solved): 90.1%; OSWorld 2.0: 70.6%; AutomationBench: 26%; Harvey's Legal Agent Benchmark (Held-out): 11.7%; HealthBench Professional: 59.8%; GDPval-AA v2: 1861; Arena Elo (Text): 1495; Arena Elo (Code): 1673. Scores are the figures published at release by Anthropic. It holds the best score among all models tracked here on Gray Swan IPI (k = 1), Gray Swan IPI (k = 10), Gray Swan IPI (k = 15), FrontierCode v1.1 (Main), Frontier-Bench v0.1, Humanity's Last Exam (no tools), Humanity's Last Exam (with tools), ARC-AGI-3, BioMysteryBench (hard), BioMysteryBench (human solved), OSWorld 2.0, AutomationBench, Harvey's Legal Agent Benchmark (Held-out), HealthBench Professional and GDPval-AA v2.
What is the context window of Claude Opus 5?
Claude Opus 5 has a context window of 1M. That is the maximum amount of input plus output the model can hold in a single request.
Is Claude Opus 5 open source?
No. Claude Opus 5 is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.
What came before and after Claude Opus 5?
Anthropic's previous tracked release was Claude Sonnet 5 on Jun 30 2026, 24 days earlier. It is the most recent Anthropic model tracked on AI Release Tracker.

Compare Claude Opus 5 with