GPT-5.4 mini
GPT-5.4 mini is an AI model released by OpenAI on Tuesday, Mar 17 2026, 12 days after GPT-5.4-Pro. It has a 400k token context window. Benchmark results (shown below) cover BullshitBench v2, Supabase Evals, BU Bench, and Arena Elo (Code).
Benchmarks
Nonsense detection
BullshitBench v2Given a confidently-worded but nonsensical prompt, does the AI spot that it makes no sense and push back — instead of playing along and inventing an answer? The score is how often it clearly called out the nonsense. Higher is better.
32%
#36 of 67Best: Claude Opus 4.8 · 95%
Supabase coding
Supabase EvalsSupabase's own open benchmark: a coding agent is dropped into a real Supabase project and asked to do real work — set up a schema, fix a broken security policy, debug an Edge Function — and every run is checked against a live Supabase stack. This is the headline number, where the agent has Supabase's own skills loaded, as most people building on Supabase would. The score is the share of scenarios it got right. Higher is better.
78.9%
with skills
73.7%
no skills
with skills
#5 of 5Best: GPT-5.6 Sol · 100%
no skills
#5 of 5Best: Kimi K3 · 100%
Browser agent
BU BenchCan the AI drive a real web browser to finish tasks — clicking, filling forms, and navigating sites the way a person would? Run by Browser Use on their BU Bench task set. Higher is better.
36%
#7 of 7Best: Claude Opus 4.8 · 74%
Community preference (code)
Arena Elo (Code)Like the text arena, but people vote on which AI writes better code. The votes become a chess-style Elo rating on arena.ai. Higher is better.
1398
#33 of 43Best: Kimi K3 · 1679
GPT-5.4 mini — frequently asked questions
- When was GPT-5.4 mini released?
- GPT-5.4 mini was released by OpenAI on Tuesday, Mar 17 2026.
- Who made GPT-5.4 mini?
- GPT-5.4 mini was built by OpenAI. Creators of ChatGPT and the GPT series of models. Pioneered large-scale language model research.
- What benchmark scores did GPT-5.4 mini get?
- GPT-5.4 mini reports 5 tracked benchmark scores — BullshitBench v2: 32%; Supabase Evals (with skills): 78.9%; Supabase Evals (no skills): 73.7%; BU Bench: 36%; Arena Elo (Code): 1398. Scores are the figures published at release by OpenAI.
- What is the context window of GPT-5.4 mini?
- GPT-5.4 mini has a context window of 400k. That is the maximum amount of input plus output the model can hold in a single request.
- Is GPT-5.4 mini open source?
- No. GPT-5.4 mini is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.
- What came before and after GPT-5.4 mini?
- OpenAI's previous tracked release was GPT-5.4-Pro on Mar 5 2026, 12 days earlier. It was followed by GPT-5.4 nano on Mar 17 2026.