Llama Nemotron Ultra 253Bvsgpt-oss-120b

Llama Nemotron Ultra 253B
gpt-oss-120b
Specifications
Parameters
253B117B
Context window
128k
API pricing
Cheapest input
$0.03AkashML
Cheapest output
$0.17AkashML
Benchmarks
GPQA Diamond
76%80.1%
Benchmarks
BullshitBench v2
11%
SWE-Bench Verified
62.4%
LiveCodeBench
66.3%
Humanity's Last Exam · no tools
14.9%
Humanity's Last Exam · with tools
19%
MMLU
90%
Overview
CompanyNVIDIAOpenAI
Release dateApr 8 2025Aug 5 2025
AccessOpen WeightOpen Weight

Other comparisons

Llama Nemotron Ultra 253BvsClaude Opus 5gpt-oss-120bvsClaude Opus 5Llama Nemotron Ultra 253BvsGemini 3.7 Flashgpt-oss-120bvsGemini 3.7 FlashLlama Nemotron Ultra 253BvsMuse Glimmergpt-oss-120bvsMuse GlimmerLlama Nemotron Ultra 253BvsGrok 4.6gpt-oss-120bvsGrok 4.6Llama Nemotron Ultra 253BvsDeepSeek-V4-Pro-0813gpt-oss-120bvsDeepSeek-V4-Pro-0813Llama Nemotron Ultra 253BvsMistral Medium 3.5gpt-oss-120bvsMistral Medium 3.5

Frequently asked questions

gpt-oss-120b leads Llama Nemotron Ultra 253B on 1 of the 1 benchmark they both report (GPQA Diamond). Llama Nemotron Ultra 253B shipped 119 days before gpt-oss-120b, so benchmark comparisons should account for the intervening progress.

Llama Nemotron Ultra 253B has 253B parameters, while gpt-oss-120b has 117B.

On GPQA Diamond, gpt-oss-120b leads at 80.1% vs Llama Nemotron Ultra 253B at 76%.

Llama Nemotron Ultra 253B was released by NVIDIA on Apr 8 2025.

gpt-oss-120b was released by OpenAI on Aug 5 2025.

gpt-oss-120b leads on GPQA Diamond — Llama Nemotron Ultra 253B 76% vs gpt-oss-120b 80.1%.