# GPT-5.6 Terra

GPT-5.6 Terra is an AI model released by OpenAI on Jun 26 2026. At release it scored 53% on BullshitBench v2, 1526 on Arena Elo (Code) and 1467 on Arena Elo (Text).

## Facts

| Field | Value |
| --- | --- |
| Model | GPT-5.6 Terra |
| Developer | OpenAI |
| Release date | Friday, Jun 26 2026 |
| Licensing | Proprietary |

## Benchmark scores published at release

| Benchmark | Score | What it measures |
| --- | --- | --- |
| BullshitBench v2 | 53% | Given a confidently-worded but nonsensical prompt, does the AI spot that it makes no sense and push back — instead of playing along and inventing an answer? The score is how often it clearly called out the nonsense. Higher is better. |
| Gray Swan IPI (k = 1) | 5.4% | Attackers hide malicious instructions inside content the AI reads — a web page, an email, a document — and try to hijack what it does. Gray Swan's indirect prompt injection benchmark measures how often such an attack succeeds when the attacker gets a single try. Lower is better. |
| Gray Swan IPI (k = 10) | 26% | Attackers hide malicious instructions inside content the AI reads — a web page, an email, a document — and try to hijack what it does. This variant gives the attacker 10 tries and counts an attack as successful if any of them works. Lower is better. |
| Gray Swan IPI (k = 15) | 30.4% | Attackers hide malicious instructions inside content the AI reads — a web page, an email, a document — and try to hijack what it does. This variant gives the attacker 15 tries and counts an attack as successful if any of them works. Lower is better. |
| CursorBench v3.2 | 64.9% | Cursor's own test of harder, real-world coding tasks inside a code editor, on the refreshed v3.2 task set. Scores aren't comparable with v3.1. Higher is better. |
| Frontier-Bench v0.1 | 20.8% | A hard, ever-evolving set of real computer tasks — coding, system administration, data work, and more — that an AI agent has to complete on its own. Run by the Harbor / Laude Institute team as the successor to Terminal-Bench (v0.1 is the first release of the task set). The score is the share of tasks solved. Higher is better. |
| Terminal-Bench 2.1 | 84.3% | Can the AI work in a command-line terminal — running commands and finishing technical setup tasks the way a developer would? Higher is better. |
| Arena Elo (Text) | 1467 | Real people chat with two anonymous AIs side by side and vote for the answer they prefer. Votes become a chess-style Elo rating on arena.ai — it measures which AI people actually like, not test scores. Higher is better. |
| Arena Elo (Code) | 1526 | Like the text arena, but people vote on which AI writes better code. The votes become a chess-style Elo rating on arena.ai. Higher is better. |

## Questions and answers

### When was GPT-5.6 Terra released?

GPT-5.6 Terra was released by OpenAI on Friday, Jun 26 2026.

### Who made GPT-5.6 Terra?

GPT-5.6 Terra was built by OpenAI. Creators of ChatGPT and the GPT series of models. Pioneered large-scale language model research.

### What benchmark scores did GPT-5.6 Terra get?

GPT-5.6 Terra reports 9 tracked benchmark scores — BullshitBench v2: 53%; Gray Swan IPI (k = 1): 5.4%; Gray Swan IPI (k = 10): 26%; Gray Swan IPI (k = 15): 30.4%; CursorBench v3.2: 64.9%; Frontier-Bench v0.1: 20.8%; Terminal-Bench 2.1: 84.3%; Arena Elo (Text): 1467; Arena Elo (Code): 1526. Scores are the figures published at release by OpenAI.

### Is GPT-5.6 Terra open source?

No. GPT-5.6 Terra is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.

### What came before and after GPT-5.6 Terra?

OpenAI's previous tracked release was GPT-5.6 Sol on Jun 26 2026. It was followed by GPT-5.6 Luna on Jun 26 2026.


---

Canonical page: https://aireleasetracker.com/model/openai/gpt-5.6-terra
Full dataset: https://aireleasetracker.com/llms-full.txt · JSON: https://aireleasetracker.com/models.json
Source: AI Release Tracker (https://aireleasetracker.com). Benchmark scores are the figures published by the releasing lab at launch.
