The model provider is the Sophnet platform. Deepseek-R1-Distill-Qwen-32B is a knowledge-distilled large language model based on Qwen 2.5 32B and trained using outputs from DeepSeek R1. DeepSeek-R1 addresses issues such as infinite repetition, poor readability, and language mixing by introducing cold-start data before reinforcement learning. DeepSeek-R1’s performance in mathematics, programming, and reasoning tasks is comparable to OpenAI-o1. To support the research community, we have open-sourced DeepSeek-R1-Zero, DeepSeek-R1, and six dense models based on Llama and Qwen. DeepSeek-R1-Distill-Qwen-32B outperforms OpenAI-o1-mini on multiple benchmark tests, setting new state-of-the-art results for dense models.
← Models
DeepSeek R1 Distill Qwen 32B
DeepSeek R1 Distill Qwen 32B Compare
Compare pricing, specifications, performance, and benchmarks for up to four models.
+ Add model · 1/4
DeepSeek R1 Distill Qwen 32B
DeepSeek · → text
Input$0.28 /M
Output$0.84 /M
Pick a second model to start comparing.
Popular comparisons
Related model match-ups readers also look at.