Deepseek R1 Distill Llama 8b — Retired model
DeepSeek R1 Distill Llama 8B is a distilled large language model based on Llama-3.1-8B-Instruct, using outputs from DeepSeek R1. The model combines advanced distillation techniques to achieve high performance across multiple benchmarks, including: - AIME 2024 pass@1: 50.4 - MATH-500 pass@1: 89.1 - CodeForces Rating: 1205 The model leverages fine-tuning from DeepSeek R1's outputs, enabling competitive performance comparable to larger frontier models. Hugging Face: - [Llama-3.1-8B](https://huggingface.co/meta-llama/Llama-3.1-8B) - [DeepSeek-R1-Distill-Llama-8B](https://huggingface.co/deepseek-ai/DeepSeek-R1-Distill-Llama-8B) |
Technical specifications
Independent benchmarks
Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.
Top 5 — CodingCoding Index leaderboard
openai/gpt-5-6-sol-xhigh78$5 / $30openai/gpt-5-6-sol-high77$5 / $30anthropic/claude-opus-5-xhigh77$5 / $25Frequently Asked Questions
Can I still use the retired Deepseek R1 Distill Llama 8b model?
No. Deepseek R1 Distill Llama 8b has been retired and is retained here only as a historical reference. See the active alternatives linked on this page.
Relay integrity checks
No integrity-check data is available for this model yet.
BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.
View the full integrity report →Related models and links
More from this provider: DeepSeek