Claude 2.1vsgpt-oss-120b

Claude 2.1
gpt-oss-120b
Specifications
Parameters
117B
Context window
200k128k
API pricing
Cheapest input
$0.03AkashML
Cheapest output
$0.17AkashML
Benchmarks
BullshitBench v2
12%
SWE-Bench Verified
62.4%
Humanity's Last Exam · no tools
14.9%
Humanity's Last Exam · with tools
19%
GPQA Diamond
80.1%
MMLU
90%
Overview
CompanyAnthropicOpenAI
Release dateNov 21 2023Aug 5 2025
AccessProprietaryOpen Weight

Other comparisons

Claude 2.1vsGemini 3.8 Flashgpt-oss-120bvsGemini 3.8 FlashClaude 2.1vsMuse Spark 1.3gpt-oss-120bvsMuse Spark 1.3Claude 2.1vsGrok 4.6gpt-oss-120bvsGrok 4.6Claude 2.1vsDeepSeek-V4.1-Flashgpt-oss-120bvsDeepSeek-V4.1-FlashClaude 2.1vsMistral Medium 3.5gpt-oss-120bvsMistral Medium 3.5Claude 2.1vsKimi K3gpt-oss-120bvsKimi K3

Frequently asked questions

Claude 2.1 and gpt-oss-120b don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Claude 2.1 shipped 623 days before gpt-oss-120b, so benchmark comparisons should account for the intervening progress.

Context windows are 200k (Claude 2.1) vs 128k (gpt-oss-120b). Claude 2.1 is proprietary, while gpt-oss-120b is open weight.

Direct benchmark comparisons are unavailable — Claude 2.1 and gpt-oss-120b don't publish scores on any of the same benchmarks.