DeepSeek: R1 Distill Llama 70B
deepseek/deepseek-r1-distill-llama-70b
DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...
Prompt $/1M
$0.800
Completion $/1M
$0.800
Context
128,000 tok
Best provider
35 tok/s · DeepInfra
Specs
- Model id
- deepseek/deepseek-r1-distill-llama-70b
- Display name
- DeepSeek: R1 Distill Llama 70B
- Modality
- text->text
- Tokenizer
- Llama3
- Instruct type
- deepseek-r1
- Context length
- 128,000
- Max completion
- 8,192
- Prompt $/1M
- $0.800
- Completion $/1M
- $0.800
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptyes
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| DeepInfra | 35 |
| Artificial Analysis | 32 |
| Novitafastest | 23 |
| deepseekfastest | — |