Benchmarks

Meta: Llama 3.3 70B Instruct

meta-llama/llama-3.3-70b-instruct

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

Prompt $/1M
$0.130
Completion $/1M
$0.400
Context
131,072 tok
Best provider
83 tok/s · Google

Specs

Model id
meta-llama/llama-3.3-70b-instruct
Display name
Meta: Llama 3.3 70B Instruct
Modality
text->text
Tokenizer
Llama3
Instruct type
llama3
Context length
131,072
Max completion
128,000
Prompt $/1M
$0.130
Completion $/1M
$0.400
Free tier
no

Capabilities

  • Function callingno
  • JSON modeno
  • System promptyes
  • Visionno
  • Moderatedno

Benchmark scores

Each axis is one benchmark, normalized to the best model in the field.

No benchmark data for this model.

Speed by provider

Providertok/s
Googlefastest83
Artificial Analysis83
Friendli62
WandB42
SambaNova41
Cloudflare25
Parasail24
Together21
Novita19
Nebius17
DeepInfra13
AkashML12
Groq12
Inceptron7
Meta: Llama 3.3 70B Instruct - Benchmarks, Pricing and Speed | TryAii