# GPT-5.4 mini

GPT-5.4 mini is an AI model released by OpenAI on Mar 17 2026. It has a 400k context window. At release it scored 32% on BullshitBench v2, 1398 on Arena Elo (Code) and 81.8% on Supabase Evals (with skills).

## Facts

| Field | Value |
| --- | --- |
| Model | GPT-5.4 mini |
| Developer | OpenAI |
| Release date | Tuesday, Mar 17 2026 |
| Licensing | Proprietary |
| Context window | 400k |

## API pricing

All rates in USD per 1,000,000 tokens, pay-as-you-go.

| Tier | Input | Cached input | Output |
| --- | --- | --- | --- |
| All documented contexts | $0.75 | $0.075 | $4.50 |

Verified August 18, 2026 against the first-party source: https://developers.openai.com/api/docs/models/gpt-5.4-mini

## Benchmark scores published at release

| Benchmark | Score | What it measures |
| --- | --- | --- |
| BullshitBench v2 | 32% | Given a confidently-worded but nonsensical prompt, does the AI spot that it makes no sense and push back — instead of playing along and inventing an answer? The score is how often it clearly called out the nonsense. Higher is better. |
| Supabase Evals (with skills) | 81.8% | Supabase's own open benchmark: a coding agent is dropped into a real Supabase project and asked to do real work — set up a schema, fix a broken security policy, debug an Edge Function — and every run is checked against a live Supabase stack. This is the headline number, where the agent has Supabase's own skills loaded, as most people building on Supabase would. The score is the share of scenarios it got right. Higher is better. |
| Supabase Evals (no skills) | 63.6% | The same Supabase scenarios, but with none of Supabase's skills loaded — so it measures what the model already knows about building on Supabase, rather than how well it follows Supabase's supplied instructions. Higher is better. |
| BU Bench | 36% | Can the AI drive a real web browser to finish tasks — clicking, filling forms, and navigating sites the way a person would? Run by Browser Use on their BU Bench task set. Higher is better. |
| Arena Elo (Code) | 1398 | Like the text arena, but people vote on which AI writes better code. The votes become a chess-style Elo rating on arena.ai. Higher is better. |

## Questions and answers

### When was GPT-5.4 mini released?

GPT-5.4 mini was released by OpenAI on Tuesday, Mar 17 2026.

### Who made GPT-5.4 mini?

GPT-5.4 mini was built by OpenAI. Creators of ChatGPT and the GPT series of models. Pioneered large-scale language model research.

### How much does GPT-5.4 mini cost?

GPT-5.4 mini costs $0.75 per million input tokens and $4.50 per million output tokens through the OpenAI API. Cached input is $0.075 per million tokens. Rates are pay-as-you-go API prices verified against OpenAI's published pricing on August 18, 2026.

### What benchmark scores did GPT-5.4 mini get?

GPT-5.4 mini reports 5 tracked benchmark scores — BullshitBench v2: 32%; Supabase Evals (with skills): 81.8%; Supabase Evals (no skills): 63.6%; BU Bench: 36%; Arena Elo (Code): 1398. Scores are the figures published at release by OpenAI.

### What is the context window of GPT-5.4 mini?

GPT-5.4 mini has a context window of 400k. That is the maximum amount of input plus output the model can hold in a single request.

### Is GPT-5.4 mini open source?

No. GPT-5.4 mini is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.

### What came before and after GPT-5.4 mini?

OpenAI's previous tracked release was GPT-5.4-Pro on Mar 5 2026, 12 days earlier. It was followed by GPT-5.4 nano on Mar 17 2026.


---

Canonical page: https://aireleasetracker.com/model/openai/gpt-5.4-mini
Full dataset: https://aireleasetracker.com/llms-full.txt · JSON: https://aireleasetracker.com/models.json
Source: AI Release Tracker (https://aireleasetracker.com). Benchmark scores are the figures published by the releasing lab at launch.
