DeepSeek: R1 Distill Qwen 32B
deepseek/deepseek-r1-distill-qwen-32b
DeepSeek R1 Distill Qwen 32B is a distilled large language model based on [Qwen 2.5 32B](https://huggingface.co/Qwen/Qwen2.5-32B), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). It outperforms OpenAI's o1-mini across various benchmarks, achieving new...
Prompt $/1M
$0.290
Completion $/1M
$0.290
Context
128,000 tok
Best provider
20 tok/s · NextBit
Specs
- Model id
- deepseek/deepseek-r1-distill-qwen-32b
- Display name
- DeepSeek: R1 Distill Qwen 32B
- Modality
- text->text
- Tokenizer
- Qwen
- Instruct type
- deepseek-r1
- Context length
- 128,000
- Max completion
- 32,768
- Prompt $/1M
- $0.290
- Completion $/1M
- $0.290
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptyes
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| NextBitfastest | 20 |
| deepseekfastest | — |
| Artificial Analysis | 0 |