Benchmarks

Qwen: Qwen2.5 VL 32B Instruct

qwen/qwen2.5-vl-32b-instruct

Qwen2.5-VL-32B is a multimodal vision-language model fine-tuned through reinforcement learning for enhanced mathematical reasoning, structured outputs, and visual problem-solving capabilities. It excels at visual analysis tasks, including object recognition, textual...

Prompt $/1M
$0.200
Completion $/1M
$0.600
Context
128,000 tok
Best provider
17 tok/s · DeepInfra

Specs

Model id
qwen/qwen2.5-vl-32b-instruct
Display name
Qwen: Qwen2.5 VL 32B Instruct
Modality
text+image->text
Tokenizer
Qwen
Instruct type
Context length
128,000
Max completion
Prompt $/1M
$0.200
Completion $/1M
$0.600
Free tier
no

Capabilities

  • Function callingno
  • JSON modeno
  • System promptno
  • Visionyes
  • Moderatedno

Benchmark scores

Each axis is one benchmark, normalized to the best model in the field.

No benchmark data for this model.

Speed by provider

Providertok/s
DeepInfrafastest17
Qwen: Qwen2.5 VL 32B Instruct - Benchmarks, Pricing and Speed | TryAii