Compare AI models

Specifications
Parameters
—117B
Context window
—128k
API pricing
Input price
$2.00—
Output price
$10.00—
Cached input price
$0.20—
Cheapest input
$2.00Amazon Bedrock$0.03AkashML
Cheapest output
$10.00Amazon Bedrock$0.17AkashML
Benchmarks
BullshitBench v2
80%12%
SWE-Bench Verified
85.2%62.4%
Humanity's Last Exam · no tools
43.2%14.9%
Humanity's Last Exam · with tools
57.4%19%
Overview
CompanyAnthropicOpenAI
Release dateJun 30 2026Aug 5 2025
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

Claude Sonnet 5 leads gpt-oss-120b on 4 of the 4 benchmarks they both report (BullshitBench v2, SWE-Bench Verified, Humanity's Last Exam). Only Claude Sonnet 5 has a verified first-party API price: $2.00 per million input tokens and $10.00 per million output tokens. No pay-as-you-go API rate is tracked for gpt-oss-120b. gpt-oss-120b shipped 329 days before Claude Sonnet 5, so benchmark comparisons should account for the intervening progress.

Claude Sonnet 5 is closed, while gpt-oss-120b is open weight.

On BullshitBench v2, Claude Sonnet 5 leads at 80% vs gpt-oss-120b at 12%. On SWE-Bench Verified, Claude Sonnet 5 leads at 85.2% vs gpt-oss-120b at 62.4%. On Humanity's Last Exam · no tools, Claude Sonnet 5 leads at 43.2% vs gpt-oss-120b at 14.9%. On Humanity's Last Exam · with tools, Claude Sonnet 5 leads at 57.4% vs gpt-oss-120b at 19%.