DeepSeek-R1-Distill-32B is a distilled model built upon Qwen2.5-32B, incorporating the best reasoning algorithms from DeepSeek-R1 and expert knowledge. It sets new records among open-source dense models across several reasoning benchmarks: AIME 2024–72.6%, MATH-500–94.3%, and others. In practice, this model is nearly on par with the distilled 70-billion-parameter version and even surpasses it in certain tests.
Technically, the model is designed for solving expert-level tasks: complex mathematical computations, code generation and analysis, scientific research, and processing long or intricate contexts. DeepSeek-R1-Distill-32B can be integrated into enterprise systems, cloud services, and platforms for automating intellectual labor. For end-user applications, it is indispensable for building expert systems, scientific assistants, platforms for automating complex business processes, and educational solutions that require thorough and well-articulated explanations in responses.
DeepSeek-R1-Distill-32B is the choice for those seeking maximum performance among open-source models without the need to move to the heaviest systems.
| Model Name | Context | Type | GPU | Status | Link |
|---|
There are no public endpoints for this model yet.
Rent your own physically dedicated instance with hourly or long-term monthly billing.
We recommend deploying private instances in the following scenarios:
| Name | GPU | TPS | Max Concurrency | |||
|---|---|---|---|---|---|---|
131,072.0 pipeline |
3 | $1.34 | 1.012 | Launch | ||
131,072.0 tensor |
4 | $1.62 | 1.601 | Launch | ||
131,072.0 pipeline |
6 | $1.65 | 1.261 | Launch | ||
131,072.0 pipeline |
3 | $2.29 | 1.100 | Launch | ||
131,072.0 |
1 | $2.37 | 1.612 | Launch | ||
131,072.0 pipeline |
3 | $2.83 | 1.097 | Launch | ||
131,072.0 tensor |
4 | $2.89 | 1.723 | Launch | ||
131,072.0 tensor |
2 | $2.93 | 1.030 | Launch | ||
131,072.0 tensor |
4 | $3.01 | 1.723 | Launch | ||
131,072.0 tensor |
4 | $3.60 | 1.718 | Launch | ||
131,072.0 |
1 | $3.83 | 1.610 | Launch | ||
131,072.0 |
1 | $4.11 | 2.010 | Launch | ||
131,072.0 tensor |
2 | $4.61 | 3.784 | Launch | ||
131,072.0 |
1 | $4.74 | 3.353 | Launch | ||
131,072.0 tensor |
2 | $9.40 | 7.266 | Launch | ||
| Name | GPU | TPS | Max Concurrency | |||
|---|---|---|---|---|---|---|
131,072.0 tensor |
4 | $1.62 | 1.161 | Launch | ||
131,072.0 |
1 | $2.37 | 1.172 | Launch | ||
131,072.0 tensor |
4 | $2.89 | 1.283 | Launch | ||
131,072.0 tensor |
4 | $3.01 | 1.283 | Launch | ||
131,072.0 tensor |
4 | $3.60 | 1.278 | Launch | ||
131,072.0 |
1 | $3.83 | 1.170 | Launch | ||
131,072.0 |
1 | $4.11 | 1.570 | Launch | ||
131,072.0 pipeline |
3 | $4.34 | 1.312 | Launch | ||
131,072.0 tensor |
2 | $4.61 | 3.344 | Launch | ||
131,072.0 |
1 | $4.74 | 2.913 | Launch | ||
131,072.0 tensor |
4 | $5.74 | 2.179 | Launch | ||
131,072.0 tensor |
2 | $9.40 | 6.826 | Launch | ||
| Name | GPU | TPS | Max Concurrency | |||
|---|---|---|---|---|---|---|
131,072.0 |
1 | $4.74 | 2.005 | Launch | ||
131,072.0 tensor |
2 | $4.93 | 2.436 | Launch | ||
131,072.0 tensor |
2 | $4.94 | 2.436 | Launch | ||
131,072.0 tensor |
4 | $5.76 | 1.271 | Launch | ||
131,072.0 pipeline |
6 | $5.84 | 1.405 | Launch | ||
131,072.0 tensor |
8 | $6.04 | 2.657 | Launch | ||
131,072.0 tensor |
8 | $7.52 | 2.647 | Launch | ||
131,072.0 tensor |
2 | $7.85 | 2.432 | Launch | ||
131,072.0 tensor |
2 | $8.17 | 3.232 | Launch | ||
131,072.0 tensor |
2 | $9.41 | 5.918 | Launch | ||
Contact our dedicated neural networks support team at support@immers.cloud or send your request to the sales department at sale@immers.cloud.