Back to Models
deepseek logo
deepseek/deepseek-r1-distill-qwen-32b
Not Available

R1 Distill Qwen 32B

DeepSeek R1 Distill Qwen 32B is a distilled large language model based on [Qwen 2.5 32B](https://huggingface.co/Qwen/Qwen2.5-32B), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). It outperforms OpenAI's o1-mini across various benchmarks, achieving new state-of-the-art results for dense models.\n\nOther benchmark results include:\n\n- AIME 2024 pass@1: 72.6\n- MATH-500 pass@1: 94.3\n- CodeForces Rating: 1691\n\nThe model leverages fine-tuning from DeepSeek R1's outputs, enabling competitive performance comparable to larger frontier models.

1/29/2025
32,768 tokens

Specifications

Modalities

Input
text
Output
text

Supported Parameters

frequency_penalty
include_reasoning
max_tokens
presence_penalty
reasoning
repetition_penalty
response_format
seed
stop
structured_outputs
temperature
top_k
top_p

Max Output Tokens

32,768

Frequently asked questions

Is R1 Distill Qwen 32B available on FastMetal?
Not at the moment. DeepSeek V4.1 Flash, from the same lab, is available on the FastMetal API today.
What is the context window of R1 Distill Qwen 32B?
32,768 tokens, shared between the prompt and the response.