Gemini 3.8 FlashvsLlama Nemotron Ultra 253B

Gemini 3.8 Flash
Llama Nemotron Ultra 253B
Specifications
Parameters
253B
Context window
1M
API pricing
Cheapest input
$0.375Google
Cheapest output
$1.875Google
Benchmarks
DeepSWE 1.1
71%
LiveCodeBench
66.3%
Terminal-Bench 4.0
19.1%
Terminal-Bench 2.1
89.4%
Humanity's Last Exam (Verified)
54.9%
BioMysteryBench · hard
56.5%
BioMysteryBench · human solved
88.8%
LAB-Bench 2
86.2%
GPQA Diamond
76%
OSWorld 2.0
59%
Finance Agent v2
61.4%
Harvey's Legal Agent Benchmark
10%
GDPval-AA v2
1545
CharXiv Reasoning
86.2%
GDP.PDF
35%
LVBench
87.1%
LVBench · agentic
87.8%
Overview
CompanyGoogleNVIDIA
Release dateSep 2 2026Apr 8 2025
AccessProprietaryOpen Weight

Other comparisons

Gemini 3.8 FlashvsClaude Fable 5.1Llama Nemotron Ultra 253BvsClaude Fable 5.1Gemini 3.8 FlashvsGPT-6 AstraLlama Nemotron Ultra 253BvsGPT-6 AstraGemini 3.8 FlashvsMuse Spark 1.3Llama Nemotron Ultra 253BvsMuse Spark 1.3Gemini 3.8 FlashvsGrok 4.6Llama Nemotron Ultra 253BvsGrok 4.6Gemini 3.8 FlashvsDeepSeek-V4-Pro-0813Llama Nemotron Ultra 253BvsDeepSeek-V4-Pro-0813Gemini 3.8 FlashvsMistral Medium 3.5Llama Nemotron Ultra 253BvsMistral Medium 3.5

Frequently asked questions

Gemini 3.8 Flash and Llama Nemotron Ultra 253B don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Llama Nemotron Ultra 253B shipped 512 days before Gemini 3.8 Flash, so benchmark comparisons should account for the intervening progress.

Gemini 3.8 Flash is proprietary, while Llama Nemotron Ultra 253B is open weight.

Direct benchmark comparisons are unavailable — Gemini 3.8 Flash and Llama Nemotron Ultra 253B don't publish scores on any of the same benchmarks.

Gemini 3.8 Flash was released by Google on Sep 2 2026.

Llama Nemotron Ultra 253B was released by NVIDIA on Apr 8 2025.

Gemini 3.8 Flash is a proprietary model released by Google. Llama Nemotron Ultra 253B is an open weight model released by NVIDIA.