Claude Opus 4.6vsGPT-5.6 Terra

Claude Opus 4.6
GPT-5.6 Terra
Benchmarks
Nonsense detection
BullshitBench v2
87%Best
53%
Prompt injection robustness
Gray Swan IPI · k = 1
5.4%
Prompt injection robustness
Gray Swan IPI · k = 10
26%
Prompt injection robustness
Gray Swan IPI · k = 15
30.4%
Coding
SWE-Bench Verified
80.8%
Agentic coding
CursorBench v3.2
64.9%
Next.js coding
Next.js Evals
75%
Agentic computer work
Frontier-Bench v0.1
20.8%
Agentic terminal coding
Terminal-Bench 2.1
84.3%
Web browsing
BrowseComp
83.7%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
53%
Science
GPQA Diamond
91.3%
Community preference
Arena Elo (Text)
1504Best
1467
Community preference (code)
Arena Elo (Code)
1543Best
1526
Overview
CompanyAnthropicOpenAI
Release dateFeb 5 2026Jun 26 2026
AccessProprietaryProprietary

Which is better: Claude Opus 4.6 or GPT-5.6 Terra?

Claude Opus 4.6 leads GPT-5.6 Terra on 3 of the 3 benchmarks they both report (BullshitBench v2, Arena Elo (Text), Arena Elo (Code)). Claude Opus 4.6 shipped 141 days before GPT-5.6 Terra, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Claude Opus 4.6 leads at 87% vs GPT-5.6 Terra at 53%. On Arena Elo (Text), Claude Opus 4.6 leads at 1504 vs GPT-5.6 Terra at 1467. On Arena Elo (Code), Claude Opus 4.6 leads at 1543 vs GPT-5.6 Terra at 1526.

Frequently asked questions

Claude Opus 4.6 was released by Anthropic on Feb 5 2026.

GPT-5.6 Terra was released by OpenAI on Jun 26 2026.

Other comparisons