Compare AI models

Specifications
Parameters
120B117B
Context window
1M128k
API pricing
Cheapest input
$0.08DekaLLM$0.03AkashML
Cheapest output
$0.40DeepInfra$0.15Venice
Benchmarks
BullshitBench v2
54%12%
SWE-Bench Verified
60.5%62.4%
GPQA Diamond
79.2%80.1%
Overview
CompanyNVIDIAOpenAI
Release dateMar 11 2026Aug 5 2025
AccessOpen SourceOpen Weight
Model detailsView modelView model

Frequently asked questions

gpt-oss-120b leads Nemotron 3 Super on 2 of the 3 benchmarks they both report (BullshitBench v2, SWE-Bench Verified, GPQA Diamond). gpt-oss-120b shipped 218 days before Nemotron 3 Super, so benchmark comparisons should account for the intervening progress.

Nemotron 3 Super has 120B parameters, while gpt-oss-120b has 117B. Context windows are 1M (Nemotron 3 Super) vs 128k (gpt-oss-120b). Nemotron 3 Super is open source, while gpt-oss-120b is open weight.

On BullshitBench v2, Nemotron 3 Super leads at 54% vs gpt-oss-120b at 12%. On SWE-Bench Verified, gpt-oss-120b leads at 62.4% vs Nemotron 3 Super at 60.5%. On GPQA Diamond, gpt-oss-120b leads at 80.1% vs Nemotron 3 Super at 79.2%.