Compare AI models

Specifications
Parameters
120B—
Context window
1M128k
API pricing
Input price
—$1.10
Output price
—$4.40
Cached input price
—$0.55
Cheapest input
$0.08DekaLLM—
Cheapest output
$0.40DeepInfra—
Benchmarks
GPQA Diamond
79.2%60%
Overview
CompanyNVIDIAOpenAI
Release dateMar 11 2026Sep 12 2024
AccessOpen SourceClosed
Model detailsView modelView model

Frequently asked questions

Nemotron 3 Super leads o1-mini on 1 of the 1 benchmark they both report (GPQA Diamond). Only o1-mini has a verified first-party API price: $1.10 per million input tokens and $4.40 per million output tokens. No pay-as-you-go API rate is tracked for Nemotron 3 Super. o1-mini shipped 545 days before Nemotron 3 Super, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Nemotron 3 Super) vs 128k (o1-mini). Nemotron 3 Super is open source, while o1-mini is closed.

On GPQA Diamond, Nemotron 3 Super leads at 79.2% vs o1-mini at 60%.