Compare AI models

Specifications
Parameters
—253B
API pricing
Input price
$1.25—
Output price
$10.00—
Cached input price
$0.125—
Cheapest input
$0.625Google AI Studio—
Cheapest output
$5.00Google AI Studio—
Benchmarks
GPQA Diamond
86.4%76%
Overview
CompanyGoogleNVIDIA
Release dateMar 25 2025Apr 8 2025
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

Gemini 2.5 Pro leads Llama Nemotron Ultra 253B on 1 of the 1 benchmark they both report (GPQA Diamond). Only Gemini 2.5 Pro has a verified first-party API price: $1.25 per million input tokens and $10.00 per million output tokens. No pay-as-you-go API rate is tracked for Llama Nemotron Ultra 253B. Gemini 2.5 Pro shipped 14 days before Llama Nemotron Ultra 253B, so benchmark comparisons should account for the intervening progress.

Gemini 2.5 Pro is closed, while Llama Nemotron Ultra 253B is open weight.

On GPQA Diamond, Gemini 2.5 Pro leads at 86.4% vs Llama Nemotron Ultra 253B at 76%.