Compare AI models

Specifications
Parameters
30B—
Context window
1M128k
API pricing
Input price
—$15.00
Output price
—$60.00
Cached input price
—$7.50
Cheapest input
$0.045Wafer—
Cheapest output
$0.13Wafer—
Benchmarks
GPQA Diamond
75.4%73.3%
Overview
CompanyNVIDIAOpenAI
Release dateAug 11 2026Sep 12 2024
AccessOpen SourceClosed
Model detailsView modelView model

Frequently asked questions

Nemotron 3.5 Lightning leads o1-preview on 1 of the 1 benchmark they both report (GPQA Diamond). Only o1-preview has a verified first-party API price: $15.00 per million input tokens and $60.00 per million output tokens. No pay-as-you-go API rate is tracked for Nemotron 3.5 Lightning. o1-preview shipped 698 days before Nemotron 3.5 Lightning, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Nemotron 3.5 Lightning) vs 128k (o1-preview). Nemotron 3.5 Lightning is open source, while o1-preview is closed.

On GPQA Diamond, Nemotron 3.5 Lightning leads at 75.4% vs o1-preview at 73.3%.