Compare AI models

Specifications
Parameters
550B—
Context window
—128k
API pricing
Input price
—$0.15
Output price
—$0.60
Cached input price
—$0.075
Cheapest input
$0.50DeepInfra$0.15Azure
Cheapest output
$2.20DeepInfra$0.60Azure
Benchmarks
BullshitBench v2
49%2%
GPQA Diamond
86.7%40.2%
Overview
CompanyNVIDIAOpenAI
Release dateJun 4 2026Jul 18 2024
AccessOpen SourceClosed
Model detailsView modelView model

Frequently asked questions

Nemotron 3 Ultra leads GPT-4o mini on 2 of the 2 benchmarks they both report (BullshitBench v2, GPQA Diamond). Only GPT-4o mini has a verified first-party API price: $0.15 per million input tokens and $0.60 per million output tokens. No pay-as-you-go API rate is tracked for Nemotron 3 Ultra. GPT-4o mini shipped 686 days before Nemotron 3 Ultra, so benchmark comparisons should account for the intervening progress.

Nemotron 3 Ultra is open source, while GPT-4o mini is closed.

On BullshitBench v2, Nemotron 3 Ultra leads at 49% vs GPT-4o mini at 2%. On GPQA Diamond, Nemotron 3 Ultra leads at 86.7% vs GPT-4o mini at 40.2%.