# GPT-6.1 Sol

GPT-6.1 Sol is an AI model released by OpenAI on Sep 29 2026. It has a 1.05M context window. Tracked results include 65% on BullshitBench v2, 75.2% on DeepSWE 1.1 and 57.1% on Terminal-Bench-Science 0.1.

## Facts

| Field | Value |
| --- | --- |
| Model | GPT-6.1 Sol |
| Developer | OpenAI |
| Release date | Tuesday, Sep 29 2026 |
| Licensing | Closed |
| Context window | 1.05M |

## API pricing

All rates in USD per 1,000,000 tokens, pay-as-you-go.

| Tier | Input | Cached input | Output |
| --- | --- | --- | --- |
| Short context; long-context pricing unavailable | $2.00 | $0.10 | $10.00 |

Verified September 29, 2026 against the first-party source: https://openai.com/index/introducing-gpt-6-1-sol/

## Tracked benchmark scores

| Benchmark | Score | Source | What it measures |
| --- | --- | --- | --- |
| BullshitBench v2 | 65% | [BullshitBench](https://github.com/petergpt/bullshit-benchmark) | Given a confidently-worded but nonsensical prompt, does the AI spot that it makes no sense and push back — instead of playing along and inventing an answer? The score is how often it clearly called out the nonsense. Higher is better. |
| Auto-review circumvention (Internal) | 0% | Lab | OpenAI's internal safety check on how often a model finds ways around its own automated review — the guardrail that inspects what it is about to do. This one counts failures, so lower is better and zero is the goal. |
| DeepSWE 1.1 | 75.2% | Lab | Artificial Analysis' independent test of deep, agentic software-engineering work — the AI has to plan and carry out substantial coding tasks end to end. (Version 1.1 of the test.) Higher is better. |
| Terminal-Bench-Science 0.1 | 57.1% | Lab | The same command-line setup as Terminal-Bench, pointed at scientific work: the AI has to drive research tooling and computational workflows through to a result, rather than administer a machine. Version 0.1 is the first release of the task set, and scores run lower than on the general board. Higher is better. |
| OSWorld 2.0 | 71.4% | Lab | Can the AI actually operate a computer — clicking, typing, and using real apps — to finish tasks on its own? Version 2.0 is a harder, refreshed task set. Higher is better. |
| AutomationBench | 36.2% | Lab | Tests whether the AI can run real multi-step business workflows — the kind of end-to-end office processes companies want to automate — from start to finish. Higher is better. |
| GDP.PDF | 32% | Lab | Real professional PDFs — filings, reports, technical documents — with questions an expert in that field would ask. Tests whether the AI reads the page as a document, layout and figures included, rather than as loose text. Higher is better. |

## About GPT-6.1 Sol

GPT-6.1 Sol, released September 29, 2026, was the first point release of OpenAI's GPT-6 generation, arriving one week after GPT-6 Sol and Luna and nearly four weeks after GPT-6 Astra opened the line. OpenAI's documentation positioned it against the flagship rather than against the Sol it replaced: "near-Astra performance at a lower cost for complex coding, computer use, and professional work", with a suggestion that developers run it beside Astra on their own tasks to judge the tradeoff between quality and cost. The list price did not move from GPT-6 Sol's: $2.00 per million input tokens and $10.00 per million output, a fifth of Astra's $10.00 and $50.00. Cached input fell to $0.10 per million tokens, half of GPT-6 Sol's cached rate.

OpenAI's launch post charted every score against cost per task. On DeepSWE 1.1 it scored 75.2% at release, 6.4 points above GPT-6 Sol's best and level with Astra at roughly a fifth of the cost. On the OSWorld 2.0 offline set it reached 71.4% at maximum effort, seven points above GPT-6 Sol and 2.1 below Astra at about a seventh of the cost per task. It scored 36.2% on AutomationBench, 32.0% on GDP.pdf and 57.1% on Terminal-Bench Science 0.1, more than double GPT-6 Sol's maximum-effort score there at $5.47 per task, against $23.80 for Astra, which had the top score in OpenAI's comparison at 68.1%. OpenAI also reported fewer factual errors on difficult prompts, 7.7% against GPT-6 Sol's 11.4% at low effort, and no attempts to bypass its automated safety reviewer.

The API surface kept the family's 1,050,000-token context window, 128,000 max output tokens, text and image input and text output, with a knowledge cutoff of April 30, 2026, ten days later than GPT-6 Sol's. The reasoning-effort control narrowed at the bottom: low, medium, high, xhigh and max were supported with medium as the default, while the none setting GPT-6 Sol had offered was dropped and minimal was not supported. Tool calling ran through the Responses API, with Chat Completions supported only without tools, and the model shipped with both US and EU data residency, Fast mode being unavailable under the EU option. It was served in the API as gpt-6.1-sol.

## Questions and answers

### When was GPT-6.1 Sol released?

GPT-6.1 Sol was released by OpenAI on Tuesday, Sep 29 2026.

### Who made GPT-6.1 Sol?

GPT-6.1 Sol was built by OpenAI. Creators of ChatGPT and the GPT series of models. Pioneered large-scale language model research.

### How much does GPT-6.1 Sol cost?

GPT-6.1 Sol costs $2.00 per million input tokens and $10.00 per million output tokens through the OpenAI API. Cached input is $0.10 per million tokens. Rates are pay-as-you-go API prices verified against OpenAI's published pricing on September 29, 2026.

### What benchmark scores did GPT-6.1 Sol get?

GPT-6.1 Sol reports 7 tracked benchmark scores — BullshitBench v2: 65%; Auto-review circumvention (Internal): 0%; DeepSWE 1.1: 75.2%; Terminal-Bench-Science 0.1: 57.1%; OSWorld 2.0: 71.4%; AutomationBench: 36.2%; GDP.PDF: 32%. Tracked scores may come from lab reports or independent benchmarks; source details accompany the benchmark data.

### What is the context window of GPT-6.1 Sol?

GPT-6.1 Sol has a context window of 1.05M. That is the maximum amount of input plus output the model can hold in a single request.

### Is GPT-6.1 Sol open source?

No. GPT-6.1 Sol is a closed model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.

### What came before and after GPT-6.1 Sol?

OpenAI's previous tracked release was GPT-6 Luna on Sep 22 2026, 7 days earlier. It is the most recent OpenAI model tracked on AI Release Tracker.


---

Canonical page: https://aireleasetracker.com/model/openai/gpt-6.1-sol
Site index: https://aireleasetracker.com/llms.txt
Source: AI Release Tracker (https://aireleasetracker.com). Most benchmark scores come from lab launch material; gathered results identify the leaderboard that published them.
