Qwen: Qwen3 VL 8B Thinking
qwen/qwen3-vl-8b-thinking
Qwen3 VL 8B Thinking is a Qwen model that isn't ranked on the TryAii Score yet — no benchmark results on file. It lists at $0.18/$2.1 per million input/output tokens with a 131K-token context. It's among the fastest models we track, at about 135 tokens/sec.
Prompt $/1M
$0.180
Completion $/1M
$2.100
Context
131,072 tok
Best provider
135 tok/s · Alibaba
Specs
- Model id
- qwen/qwen3-vl-8b-thinking
- Display name
- Qwen: Qwen3 VL 8B Thinking
- Modality
- text+image->text
- Tokenizer
- Qwen3
- Instruct type
- —
- Context length
- 131,072
- Max completion
- 32,768
- Prompt $/1M
- $0.180
- Completion $/1M
- $2.100
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionyes
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Alibabafastest | 135 |
| qwenfastest | — |