Deepseek R1 Distill Qwen 1.5b — Retired model
DeepSeek R1 Distill Qwen 1.5B is a distilled large language model based on [Qwen 2.5 Math 1.5B](https://huggingface.co/Qwen/Qwen2.5-Math-1.5B), using outputs from DeepSeek R1. It's a very small and efficient model which outperforms GPT 4o 0513 on Math Benchmarks. Other benchmark results include: - AIME 2024 pass@1: 28.9 - AIME 2024 cons@64: 52.7 - MATH-500 pass@1: 83.9 The model leverages fine-tuning from DeepSeek R1's outputs, enabling competitive performance comparable to larger frontier models.
Technical specifications
Independent benchmarks
Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.
Top 5 — CodingCoding Index leaderboard
openai/gpt-5-6-sol-xhigh78$5 / $30openai/gpt-5-6-sol-high77$5 / $30anthropic/claude-opus-5-xhigh77$5 / $25Frequently Asked Questions
Can I still use the retired Deepseek R1 Distill Qwen 1.5b model?
No. Deepseek R1 Distill Qwen 1.5b has been retired and is retained here only as a historical reference. See the active alternatives linked on this page.
Relay integrity checks
No integrity-check data is available for this model yet.
BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.
View the full integrity report →Related models and links
More from this provider: DeepSeek