Meta: Llama 3.3 70B Instruct
meta-llama/llama-3.3-70b-instruct
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
Prompt $/1M
$0.130
Completion $/1M
$0.400
Context
131,072 tok
Best provider
83 tok/s · Google
Specs
- Model id
- meta-llama/llama-3.3-70b-instruct
- Display name
- Meta: Llama 3.3 70B Instruct
- Modality
- text->text
- Tokenizer
- Llama3
- Instruct type
- llama3
- Context length
- 131,072
- Max completion
- 128,000
- Prompt $/1M
- $0.130
- Completion $/1M
- $0.400
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptyes
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Googlefastest | 83 |
| Artificial Analysis | 83 |
| Friendli | 62 |
| WandB | 42 |
| SambaNova | 41 |
| Cloudflare | 25 |
| Parasail | 24 |
| Together | 21 |
| Novita | 19 |
| Nebius | 17 |
| DeepInfra | 13 |
| AkashML | 12 |
| Groq | 12 |
| Inceptron | 7 |