Compare AI models

Specifications
Parameters
120B—
Context window
1M128k
API pricing
Input price
—$2.50
Output price
—$10.00
Cached input price
—$1.25
Cheapest input
$0.08DekaLLM$2.50Azure
Cheapest output
$0.40DeepInfra$10.00Azure
Benchmarks
BullshitBench v2
54%12%
GPQA Diamond
79.2%49.9%
Overview
CompanyNVIDIAOpenAI
Release dateMar 11 2026May 13 2024
AccessOpen SourceClosed
Model detailsView modelView model

Frequently asked questions

Nemotron 3 Super leads GPT-4o on 2 of the 2 benchmarks they both report (BullshitBench v2, GPQA Diamond). Only GPT-4o has a verified first-party API price: $2.50 per million input tokens and $10.00 per million output tokens. No pay-as-you-go API rate is tracked for Nemotron 3 Super. GPT-4o shipped 667 days before Nemotron 3 Super, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Nemotron 3 Super) vs 128k (GPT-4o). Nemotron 3 Super is open source, while GPT-4o is closed.

On BullshitBench v2, Nemotron 3 Super leads at 54% vs GPT-4o at 12%. On GPQA Diamond, Nemotron 3 Super leads at 79.2% vs GPT-4o at 49.9%.