Benchmarks

DeepSeek: R1 Distill Llama 70B

deepseek/deepseek-r1-distill-llama-70b

DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...

Prompt $/1M
$0.800
Completion $/1M
$0.800
Context
128,000 tok
Best provider
35 tok/s · DeepInfra

Specs

Model id
deepseek/deepseek-r1-distill-llama-70b
Display name
DeepSeek: R1 Distill Llama 70B
Modality
text->text
Tokenizer
Llama3
Instruct type
deepseek-r1
Context length
128,000
Max completion
8,192
Prompt $/1M
$0.800
Completion $/1M
$0.800
Free tier
no

Capabilities

  • Function callingno
  • JSON modeno
  • System promptyes
  • Visionno
  • Moderatedno

Benchmark scores

Each axis is one benchmark, normalized to the best model in the field.

No benchmark data for this model.

Speed by provider

Providertok/s
DeepInfra35
Artificial Analysis32
Novitafastest23
deepseekfastest
DeepSeek: R1 Distill Llama 70B - Benchmarks, Pricing and Speed | TryAii