Meta: Llama 3.2 11B Vision Instruct
meta-llama/llama-3.2-11b-vision-instruct
Llama 3.2 11B Vision Instruct (Meta) is a compact 11B open-weight model priced near the bottom of the board, at $0.345/$0.345 per million input/output tokens, served by 1 independent host. It last ranked #234 of 235 on the TryAii Score, but its benchmark evidence has gone stale. It's reasonably quick at roughly 52-64 tokens/sec.
Prompt $/1M
$0.345
Completion $/1M
$0.345
Context
131,072 tok
Best provider
70 tok/s · Artificial Analysis
Specs
- Model id
- meta-llama/llama-3.2-11b-vision-instruct
- Display name
- Meta: Llama 3.2 11B Vision Instruct
- Modality
- text+image->text
- Tokenizer
- Llama3
- Instruct type
- llama3
- Context length
- 131,072
- Max completion
- 16,384
- Prompt $/1M
- $0.345
- Completion $/1M
- $0.345
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptyes
- Visionyes
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Artificial Analysis | 70 |
| DeepInfrafastest | 46 |