Compare AI models

Specifications
Parameters
—550B
API pricing
Input price
$0.30—
Output price
$2.50—
Cached input price
$0.03—
Cheapest input
$0.15Google$0.50DeepInfra
Cheapest output
$1.25Google$2.20DeepInfra
Benchmarks
Terminal-Bench 2.1
54%56.4%
Overview
CompanyGoogleNVIDIA
Release dateJul 21 2026Jun 4 2026
AccessClosedOpen Source
Model detailsView modelView model

Frequently asked questions

Nemotron 3 Ultra leads Gemini 3.5 Flash-Lite on 1 of the 1 benchmark they both report (Terminal-Bench 2.1). Only Gemini 3.5 Flash-Lite has a verified first-party API price: $0.30 per million input tokens and $2.50 per million output tokens. No pay-as-you-go API rate is tracked for Nemotron 3 Ultra. Nemotron 3 Ultra shipped 47 days before Gemini 3.5 Flash-Lite, so benchmark comparisons should account for the intervening progress.

Gemini 3.5 Flash-Lite is closed, while Nemotron 3 Ultra is open source.

On Terminal-Bench 2.1, Nemotron 3 Ultra leads at 56.4% vs Gemini 3.5 Flash-Lite at 54%.