Compare AI models

Specifications
Parameters
—550B
Context window
1M—
API pricing
Input price
$0.75—
Output price
$3.75—
Cached input price
$0.075—
Cheapest input
$0.375Google$0.50DeepInfra
Cheapest output
$1.875Google$2.20DeepInfra
Benchmarks
BullshitBench v2
35%49%
Terminal-Bench 2.1
85.8%56.4%
AA Intelligence Index
5648
Overview
CompanyGoogleNVIDIA
Release dateAug 13 2026Jun 4 2026
AccessClosedOpen Source
Model detailsView modelView model

Frequently asked questions

Gemini 3.7 Flash leads Nemotron 3 Ultra on 2 of the 3 benchmarks they both report (BullshitBench v2, Terminal-Bench 2.1, AA Intelligence Index). Only Gemini 3.7 Flash has a verified first-party API price: $0.75 per million input tokens and $3.75 per million output tokens. No pay-as-you-go API rate is tracked for Nemotron 3 Ultra. Nemotron 3 Ultra shipped 70 days before Gemini 3.7 Flash, so benchmark comparisons should account for the intervening progress.

Gemini 3.7 Flash is closed, while Nemotron 3 Ultra is open source.

On BullshitBench v2, Nemotron 3 Ultra leads at 49% vs Gemini 3.7 Flash at 35%. On Terminal-Bench 2.1, Gemini 3.7 Flash leads at 85.8% vs Nemotron 3 Ultra at 56.4%. On AA Intelligence Index, Gemini 3.7 Flash leads at 56 vs Nemotron 3 Ultra at 48.