Llama 3.1 Nemotron 70BvsGLM-5.2

Llama 3.1 Nemotron 70B
GLM-5.2
Specifications
Parameters
70B744B
Context window
1M
API pricing
Input price
$1.40
Output price
$4.40
Cached input price
$0.26
Cheapest input
$0.50Sail Research
Cheapest output
$2.00Ambient
Benchmarks
BullshitBench v2
31%
SWE-Bench Pro
62.1%
DeepSWE 1.1
44%
Next.js Evals
88%
Frontier-Bench v0.1
5.1%
Terminal-Bench 2.1
81%
Humanity's Last Exam · no tools
40.5%
Humanity's Last Exam · with tools
54.7%
GPQA Diamond
91.2%
GDPval-AA v2
1514
Overview
CompanyNVIDIAZ.ai
Release dateOct 15 2024Jun 16 2026
AccessOpen WeightOpen Weight

Other comparisons

Llama 3.1 Nemotron 70BvsClaude Opus 5GLM-5.2vsClaude Opus 5Llama 3.1 Nemotron 70BvsGPT-5.6 SolGLM-5.2vsGPT-5.6 SolLlama 3.1 Nemotron 70BvsGemini 3.7 FlashGLM-5.2vsGemini 3.7 FlashLlama 3.1 Nemotron 70BvsMuse GlimmerGLM-5.2vsMuse GlimmerLlama 3.1 Nemotron 70BvsGrok 4.6GLM-5.2vsGrok 4.6Llama 3.1 Nemotron 70BvsDeepSeek-V4-Pro-0813GLM-5.2vsDeepSeek-V4-Pro-0813

Frequently asked questions

Llama 3.1 Nemotron 70B and GLM-5.2 don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Only GLM-5.2 has a verified first-party API price: $1.40 per million input tokens and $4.40 per million output tokens. No pay-as-you-go API rate is tracked for Llama 3.1 Nemotron 70B. Llama 3.1 Nemotron 70B shipped 609 days before GLM-5.2, so benchmark comparisons should account for the intervening progress.

Llama 3.1 Nemotron 70B has 70B parameters, while GLM-5.2 has 744B.

Direct benchmark comparisons are unavailable — Llama 3.1 Nemotron 70B and GLM-5.2 don't publish scores on any of the same benchmarks.

Llama 3.1 Nemotron 70B was released by NVIDIA on Oct 15 2024.

GLM-5.2 was released by Z.ai on Jun 16 2026.

Only GLM-5.2 has a verified first-party API price: $1.40 per million input tokens and $4.40 per million output tokens. No pay-as-you-go API rate is tracked for Llama 3.1 Nemotron 70B. Rates are pay-as-you-go API prices verified on August 18, 2026.