Compare AI models

Specifications
Parameters
550B—
Context window
—128k
API pricing
Input price
—$10.00
Output price
—$30.00
Cheapest input
$0.50DeepInfra$10.00OpenAI
Cheapest output
$2.20DeepInfra$30.00OpenAI
Benchmarks
GPQA Diamond
86.7%42.5%
Overview
CompanyNVIDIAOpenAI
Release dateJun 4 2026Nov 6 2023
AccessOpen SourceClosed
Model detailsView modelView model

Frequently asked questions

Nemotron 3 Ultra leads GPT-4 Turbo on 1 of the 1 benchmark they both report (GPQA Diamond). Only GPT-4 Turbo has a verified first-party API price: $10.00 per million input tokens and $30.00 per million output tokens. No pay-as-you-go API rate is tracked for Nemotron 3 Ultra. GPT-4 Turbo shipped 941 days before Nemotron 3 Ultra, so benchmark comparisons should account for the intervening progress.

Nemotron 3 Ultra is open source, while GPT-4 Turbo is closed.

On GPQA Diamond, Nemotron 3 Ultra leads at 86.7% vs GPT-4 Turbo at 42.5%.