DeepSeek: R1 Distill Qwen 32B
DeepSeek R1 Distill Qwen 32B is a distilled large language model based on [Qwen 2.5 32B](https://huggingface.co/Qwen/Qwen2.5-32B), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). It outperforms OpenAI's o1-mini across various benchmarks, achieving new state-of-the-art results for dense models. Other benchmark results include: - AIME 2024 pass@1: 72.6 - MATH-500 pass@1: 94.3 - CodeForces Rating: 1691 The model leverages fine-tuning from DeepSeek R1's outputs, enabling competitive performance comparable to larger frontier models.
Undisclosed
Parameters
33K tokens
Context Window
Proprietary
License
Jan 29, 2025
Released
๐ฐ Pricing
Input
$0.29
per 1M tokens
Output
$0.29
per 1M tokens
API Available
This model is accessible via API for integration into your applications.
โญ Related Models
Claude 4 Opus
Anthropic
Anthropic's most powerful reasoning model with extended thinking. Excels at complex analysis, multi-step math, advanced coding, and nuanced writing.
Claude 4 Sonnet
Anthropic
Balanced intelligence and speed. Strong reasoning with faster response times and lower cost than Opus.
o3
OpenAI
OpenAI's most powerful reasoning model. Uses chain-of-thought to solve complex math, science, and coding problems.