Meta: Llama 3.3 70B Instruct
meta-llama/llama-3.3-70b-instruct
Llama 3.3 70B Instruct is a 70B open-weight model, ranking #166 of 234 on the TryAii Score. Its best showing is #2 on GSM8K. It's served by 14 independent hosts from $0.1/$0.32 per million input/output tokens — about a third of what models at its level usually charge.
Prompt $/1M
$0.100
Completion $/1M
$0.320
Context
131,072 tok
Best provider
176 tok/s · Groq
Specs
- Model id
- meta-llama/llama-3.3-70b-instruct
- Display name
- Meta: Llama 3.3 70B Instruct
- Modality
- text->text
- Tokenizer
- Llama3
- Instruct type
- llama3
- Context length
- 131,072
- Max completion
- 16,384
- Prompt $/1M
- $0.100
- Completion $/1M
- $0.320
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptyes
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Groqfastest | 176 |
| Artificial Analysis | 86 |
| Friendli | 62 |
| CoreWeave | 51 |
| WandB | 40 |
| Crusoe | 38 |
| Parasail | 32 |
| SambaNova | 30 |
| Cloudflare | 28 |
| 28 | |
| Novita | 25 |
| Together | 18 |
| AkashML | 18 |
| DeepInfra | 15 |
| Nebius | 9 |
| Inceptron | 7 |