Compare AI models

Specifications
Parameters
—550B
API pricing
Input price
$0.30—
Output price
$2.50—
Cached input price
$0.03—
Cheapest input
$0.15Google AI Studio$0.50DeepInfra
Cheapest output
$1.25Google AI Studio$2.20DeepInfra
Benchmarks
BullshitBench v2
19%49%
Overview
CompanyGoogleNVIDIA
Release dateApr 17 2025Jun 4 2026
AccessClosedOpen Source
Model detailsView modelView model

Frequently asked questions

Nemotron 3 Ultra leads Gemini 2.5 Flash on 1 of the 1 benchmark they both report (BullshitBench v2). Only Gemini 2.5 Flash has a verified first-party API price: $0.30 per million input tokens and $2.50 per million output tokens. No pay-as-you-go API rate is tracked for Nemotron 3 Ultra. Gemini 2.5 Flash shipped 413 days before Nemotron 3 Ultra, so benchmark comparisons should account for the intervening progress.

Gemini 2.5 Flash is closed, while Nemotron 3 Ultra is open source.

On BullshitBench v2, Nemotron 3 Ultra leads at 49% vs Gemini 2.5 Flash at 19%.