Meta: Llama 3 8B Instruct
meta-llama/llama-3-8b-instruct
Llama 3 8B Instruct (Meta) is a compact 8B open-weight model priced near the bottom of the board, at $0.14/$0.14 per million input/output tokens, served by 4 independent hosts. For the money it delivers #200 of 234 on the TryAii Score. It's among the fastest models we track, at roughly 34-67 tokens/sec.
Prompt $/1M
$0.140
Completion $/1M
$0.140
Context
8,192 tok
Best provider
78 tok/s · Artificial Analysis
Specs
- Model id
- meta-llama/llama-3-8b-instruct
- Display name
- Meta: Llama 3 8B Instruct
- Modality
- text->text
- Tokenizer
- Llama3
- Instruct type
- llama3
- Context length
- 8,192
- Max completion
- —
- Prompt $/1M
- $0.140
- Completion $/1M
- $0.140
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptyes
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Artificial Analysis | 78 |
| DeepInfra | 67 |
| Novitafastest | 63 |
| Togetherfastest | 34 |
| Cloudflare | 14 |
| meta-llamafastest | — |