Benchmarks

Qwen: Qwen3 VL 8B Thinking

qwen/qwen3-vl-8b-thinking

Qwen3 VL 8B Thinking is a Qwen model that isn't ranked on the TryAii Score yet — no benchmark results on file. It lists at $0.18/$2.1 per million input/output tokens with a 131K-token context. It's among the fastest models we track, at about 135 tokens/sec.

Prompt $/1M
$0.180
Completion $/1M
$2.100
Context
131,072 tok
Best provider
135 tok/s · Alibaba

Specs

Model id
qwen/qwen3-vl-8b-thinking
Display name
Qwen: Qwen3 VL 8B Thinking
Modality
text+image->text
Tokenizer
Qwen3
Instruct type
Context length
131,072
Max completion
32,768
Prompt $/1M
$0.180
Completion $/1M
$2.100
Free tier
no

Capabilities

  • Function callingno
  • JSON modeno
  • System promptno
  • Visionyes
  • Moderatedno

Benchmark scores

Each axis is one benchmark, normalized to the best model in the field.

No benchmark data for this model.

Speed by provider

Providertok/s
Alibabafastest135
qwenfastest