Meta: Llama 3.1 8B Instruct
meta-llama/llama-3.1-8b-instruct
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...
Prompt $/1M
$0.050
Completion $/1M
$0.080
Context
131,072 tok
Best provider
304 tok/s · Cerebras
Specs
- Model id
- meta-llama/llama-3.1-8b-instruct
- Display name
- Meta: Llama 3.1 8B Instruct
- Modality
- text->text
- Tokenizer
- Llama3
- Instruct type
- llama3
- Context length
- 131,072
- Max completion
- 131,072
- Prompt $/1M
- $0.050
- Completion $/1M
- $0.080
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptyes
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Cerebrasfastest | 304 |
| Friendli | 171 |
| Artificial Analysis | 165 |
| WandBfastest | 104 |
| Groq | 76 |
| Novita | 60 |
| Nebius | 20 |
| DeepInfra | 20 |
| Cloudflare | 17 |
| SambaNova | 12 |