BazaarLinkBazaarLink
Sign in

Llama 3.1 Nemotron 70b Instruct — Retired model

Retired modelThis model has reached end of life and is no longer available through BazaarLink. Historical specifications and independent benchmarks remain below. Retired on May 12, 2026.

NVIDIA's Llama 3.1 Nemotron 70B is a language model designed for generating precise and useful responses. Leveraging Llama 3.1 70B architecture and Reinforcement Learning from Human Feedback (RLHF), it excels in automatic alignment benchmarks. This model is tailored for applications requiring high accuracy in helpfulness and response generation, suitable for diverse user queries across multiple domains. Usage of this model is subject to [Meta's Acceptable Use Policy](https://www.llama.com/llama3/use-policy/).

ProviderNVIDIA
ReleasedOctober 2024
Model IDnvidia/llama-3.1-nemotron-70b-instruct

Technical specifications

Context window
131K tokens
Reasoning
Input
Text
Output
Text

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 7
Coding
Math 11
MMLU 69
GPQA 47
This modelCategory leader
Overall intelligence
#485out of 638 models
Aggregated across academic benchmarks
Time to first token
#262out of 329 models
AA 全球 p50:8.69 秒
Intelligence
7
Mid-tier
This model
7
Category leader
53
Median
12
Math
11
This model
11
Category leader
99
Median
53
MMLU Pro
69%
This model
69%
Category leader
90%
Median
75%
GPQA
47%
This model
47%
Category leader
94%
Median
69%
LiveCodeBench
17%
This model
17%
Category leader
42%
Median
42%
HLE
5%
This model
5%
Category leader
53%
Median
7%
SciCode
23%
This model
23%
Category leader
60%
Median
34%
IFBench
31%
This model
31%
Category leader
83%
Median
44%
τ²-Bench Telecom
23%
This model
23%
Category leader
99%
Median
46%
AA-LCR
7%
This model
7%
Category leader
76%
Median
40%
Terminal-Bench Hard
5%
This model
5%
Category leader
66%
Median
14%

Top 5 — CodingCoding Index leaderboard

2Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback)?/claude-fable-5-1-xhigh81$10 / $50
3Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)?/claude-fable-5-1-high79$10 / $50
4GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$4 / $20
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →
← All Models

Frequently Asked Questions

Can I still use the retired Llama 3.1 Nemotron 70b Instruct model?

No. Llama 3.1 Nemotron 70b Instruct has been retired and is retained here only as a historical reference. See the active alternatives linked on this page.

Relay integrity checks

No integrity-check data is available for this model yet.

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

Related models and links

Qwen3.8 MaxDeepseek V3.2Glm 5.1Deepseek V4 FlashGlm 4.7Minimax M2.7
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.