BazaarLinkBazaarLink
Sign in

Llama 3.1 Nemotron 70b Instruct — Retired model

Retired modelThis model has reached end of life and is no longer available through BazaarLink. Historical specifications and independent benchmarks remain below. Retired on May 12, 2026.

NVIDIA's Llama 3.1 Nemotron 70B is a language model designed for generating precise and useful responses. Leveraging Llama 3.1 70B architecture and Reinforcement Learning from Human Feedback (RLHF), it excels in automatic alignment benchmarks. This model is tailored for applications requiring high accuracy in helpfulness and response generation, suitable for diverse user queries across multiple domains. Usage of this model is subject to [Meta's Acceptable Use Policy](https://www.llama.com/llama3/use-policy/).

ProviderNVIDIA
ReleasedOctober 2024
Model IDnvidia/llama-3.1-nemotron-70b-instruct

Technical specifications

Context window
131K tokens
Reasoning
Input
Text
Output
Text

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 8
Coding
Math 11
MMLU 69
GPQA 47
This modelCategory leader
Overall intelligence
#427out of 573 models
Aggregated across academic benchmarks
Math reasoning
#230out of 269 models
AIME / MATH-500 / SciCode
Time to first token
#543out of 586 models
AA 全球 p50:4.55 秒
Intelligence
8
Mid-tier
This model
8
Category leader
61
Median
15
Math
11
Mid-tier
This model
11
Category leader
99
Median
53
MMLU Pro
69%
This model
69%
Category leader
90%
Median
75%
GPQA
47%
This model
47%
Category leader
94%
Median
69%
LiveCodeBench
17%
This model
17%
Category leader
42%
Median
42%
HLE
5%
This model
5%
Category leader
53%
Median
7%
SciCode
23%
This model
23%
Category leader
60%
Median
33%
IFBench
31%
This model
31%
Category leader
83%
Median
44%
τ²-Bench Telecom
23%
This model
23%
Category leader
99%
Median
46%
AA-LCR
7%
This model
7%
Category leader
76%
Median
40%
Terminal-Bench Hard
5%
This model
5%
Category leader
66%
Median
14%

Top 5 — CodingCoding Index leaderboard

1GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$5 / $30
4GPT-5.6 Sol (high)openai/gpt-5-6-sol-high77$5 / $30
5Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)anthropic/claude-opus-5-xhigh77$5 / $25
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →
← All Models

Frequently Asked Questions

Can I still use the retired Llama 3.1 Nemotron 70b Instruct model?

No. Llama 3.1 Nemotron 70b Instruct has been retired and is retained here only as a historical reference. See the active alternatives linked on this page.

Relay integrity checks

No integrity-check data is available for this model yet.

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: NVIDIA

Nemotron 3 Super 120b A12bNemotron 3 Nano 30b A3bGlm 4.7Glm 5Glm 5.2Deepseek V3.2
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.