Qwen: Qwen2.5 VL 32B Instruct
qwen/qwen2.5-vl-32b-instruct
Qwen2.5-VL-32B is a multimodal vision-language model fine-tuned through reinforcement learning for enhanced mathematical reasoning, structured outputs, and visual problem-solving capabilities. It excels at visual analysis tasks, including object recognition, textual...
Prompt $/1M
$0.200
Completion $/1M
$0.600
Context
128,000 tok
Best provider
17 tok/s · DeepInfra
Specs
- Model id
- qwen/qwen2.5-vl-32b-instruct
- Display name
- Qwen: Qwen2.5 VL 32B Instruct
- Modality
- text+image->text
- Tokenizer
- Qwen
- Instruct type
- —
- Context length
- 128,000
- Max completion
- —
- Prompt $/1M
- $0.200
- Completion $/1M
- $0.600
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionyes
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| DeepInfrafastest | 17 |