gpt-oss-20bOpen Weight
Released
gpt-oss-20b is an AI model released by OpenAI on Tuesday, Aug 5 2025, the same day as gpt-oss-120b. It is an open-weight model — the trained weights are available to download and run. It is a 21B parameter model with a 128k token context window. Benchmark results (shown below) cover SWE-Bench Verified, Humanity's Last Exam, GPQA Diamond, and MMLU.
Get OpenAI releases and AI news.
One email, every other Monday.
Available from
| AkashML | — | $0.02 | $0.10 | 128K | fp4 |
|---|---|---|---|---|---|
| Darkbloom | — | $0.02 | $0.10 | 128K | fp8 |
| CoreWeave | — | $0.03 | $0.13 | 128K | fp4 |
| DekaLLM | — | $0.029 | $0.14 | 128K | bf16 |
| DeepInfra | — | $0.03 | $0.14 | 128K | bf16 |
| Parasail | — | $0.03 | $0.15 | 128K | fp4 |
| Novita | — | $0.04 | $0.15 | 128K | fp4 |
| Phala | — | $0.04 | $0.15 | 128K | — |
| Amazon Bedrock | — | $0.07 | $0.15 | 128K | — |
| Amazon Bedrock | eu-west-1 | $0.07 | $0.15 | 128K | — |
| SiliconFlow | — | $0.04 | $0.18 | 128K | fp8 |
| Together | — | $0.05 | $0.20 | 128K | — |
| us-central1 | $0.07 | $0.25 | 128K | — | |
| Groq | — | $0.075 | $0.30 | 128K | — |
Benchmarks
Coding
SWE-Bench VerifiedCoding — Real coding tasks pulled from open-source projects — the AI has to find and fix actual bugs. A human-checked version of the original SWE-Bench. Higher is better.
60.7%
#46 of 57Best: Claude Opus 5 · 96%
Reasoning & science
Humanity's Last ExamMultidisciplinary reasoning — Humanity's Last Exam — extremely hard expert questions across many subjects, written so you can't just look up the answer. “No tools” means the AI answers on its own. Higher is better.
10.9%
no tools
17.3%
with tools
no tools
#22 of 22Best: Claude Fable 5.1 · 60.9%
with tools
#42 of 42Best: Claude Fable 5.1 · 65%
GPQA DiamondScience — Graduate-level science questions in biology, physics, and chemistry — hard enough that subject-matter PhDs score around 65%. Higher is better.
71.5%
#48 of 59Best: GPT-6 Astra · 96%
MMLUGeneral knowledge — A 57-subject multiple-choice exam — history, law, medicine, maths — that was the standard measure of how much a model knows from 2020 until roughly 2024, when frontier scores crowded into the high 80s and labs moved on to harder tests. The scores here were published years apart under different testing setups, so read them as a historical record rather than a like-for-like ranking. Higher is better.
85.3%
#3 of 4Best: gpt-oss-120b · 90%
Compare gpt-oss-20b with
Suggested comparisons
gpt-oss-20bvsGPT-6 Astragpt-oss-20bvsClaude Fable 5.1gpt-oss-20bvsGemini 3.8 Flashgpt-oss-20bvsMuse Spark 1.3gpt-oss-20bvsGrok 4.6gpt-oss-20bvsDeepSeek-V4.1-Flashgpt-oss-20bvsMistral Medium 3.5gpt-oss-20bvsKimi K3gpt-oss-20bvsGLM-5.3-Flashgpt-oss-20bvsQwen3.8-Max-0902gpt-oss-20bvsNemotron 3.5 Lightning
Frequently asked questions
gpt-oss-20b was released by OpenAI on Tuesday, Aug 5 2025.
All OpenAI releases
44 tracked2026
14 releasesGPT-6 Astra
Sep 3 2026
GPT-5.6-Cyber
Aug 10 2026
GPT-5.6 Sol
Jun 26 2026
GPT-5.6 Terra
Jun 26 2026
GPT-5.6 Luna
Jun 26 2026
GPT-5.5-Cyber
Jun 22 2026
GPT-5.5
Apr 23 2026
GPT-5.5-Pro
Apr 23 2026
GPT-5.4 mini
Mar 17 2026
GPT-5.4 nano
Mar 17 2026
GPT-5.4
Mar 5 2026
GPT-5.4-Pro
Mar 5 2026
GPT-5.3-Codex-Spark
Feb 12 2026
GPT-5.3-Codex
Feb 5 2026
2025
22 releasesGPT-5.2
Dec 11 2025
GPT-5.1-Codex-Max
Nov 19 2025
GPT-5.1
Nov 12 2025
GPT-5-Codex-Mini
Nov 7 2025
GPT-5-Codex
Sep 15 2025
GPT-5
Aug 7 2025
GPT-5 mini
Aug 7 2025
GPT-5 nano
Aug 7 2025
GPT-5 Pro
Aug 7 2025
gpt-oss-120b
Aug 5 2025
gpt-oss-20b
Aug 5 2025
o3-pro
Jun 10 2025
o3
Apr 16 2025
o4-mini
Apr 16 2025
o4-mini-high
Apr 16 2025
GPT-4.1
Apr 14 2025
GPT-4.1 mini
Apr 14 2025
GPT-4.1 nano
Apr 14 2025
o1-pro
Mar 19 2025
GPT-4.5
Feb 27 2025
o3-mini
Jan 31 2025
o3-mini-high
Jan 31 2025