Llama Nemotron (Nano/Super)Open Weight
Released
Llama Nemotron (Nano/Super) is an AI model released by NVIDIA on Tuesday, Mar 18 2025, 154 days after Llama 3.1 Nemotron 70B. It is an open-weight model — the trained weights are available to download and run. It was released in 8B and 49B parameter sizes.
Get NVIDIA releases and AI news.
Every other Monday
Compare Llama Nemotron (Nano/Super) with
Suggested comparisons
Llama Nemotron (Nano/Super)vsNemotron 3.5 LightningLlama Nemotron (Nano/Super)vsGPT-6.1 SolLlama Nemotron (Nano/Super)vsClaude Haiku 5.5Llama Nemotron (Nano/Super)vsGemini 4 ArgonLlama Nemotron (Nano/Super)vsMuse Spark 1.3Llama Nemotron (Nano/Super)vsGrok 4.7Llama Nemotron (Nano/Super)vsDeepSeek-V4.1-FlashLlama Nemotron (Nano/Super)vsMistral Large 4Llama Nemotron (Nano/Super)vsKimi K3Llama Nemotron (Nano/Super)vsGLM-5.3-FlashLlama Nemotron (Nano/Super)vsQwen3.8-Max-0902About
The Llama Nemotron family, launched March 18, 2025 at NVIDIA's GTC conference, was the company's entry into the reasoning-model wave that DeepSeek-R1 had set off that January. Nano (8B, distilled from Llama 3.1) and Super (49B, compressed from Llama 3.3 70B by neural architecture search) shipped first as open weights, with an Ultra tier announced for multi-GPU servers. Each model carried a switchable reasoning mode — step-by-step thinking toggled on or off through the system prompt, the same hybrid idea Claude 3.7 Sonnet had introduced three weeks earlier.
NVIDIA positioned the family explicitly as infrastructure for agentic AI platforms rather than as a chatbot, claiming up to 20% accuracy gains over the Llama bases and large inference-speed multiples from the compression work. The releases came with most of the post-training data and the tools to reproduce the recipe, and the Nano/Super/Ultra size ladder became the naming scheme the Nemotron line kept from then on.
Frequently asked questions
Llama Nemotron (Nano/Super) was released by NVIDIA on Tuesday, Mar 18 2025.