Qwen: Qwen3 VL 30B A3B Instruct
qwen/qwen3-vl-30b-a3b-instruct
Qwen3 VL 30B A3B Instruct is a 30B open-weight model (3B active per token), ranking #161 of 235 on the TryAii Score. It's served by 7 independent hosts from $0.15/$0.6 per million input/output tokens — a little under the going rate for its score class. It's on the slow side at roughly 10-22 tokens/sec.
Prompt $/1M
$0.150
Completion $/1M
$0.600
Context
262,144 tok
Best provider
51 tok/s · AtlasCloud
Specs
- Model id
- qwen/qwen3-vl-30b-a3b-instruct
- Display name
- Qwen: Qwen3 VL 30B A3B Instruct
- Modality
- text+image->text
- Tokenizer
- Qwen3
- Instruct type
- —
- Context length
- 262,144
- Max completion
- 16,384
- Prompt $/1M
- $0.150
- Completion $/1M
- $0.600
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionyes
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| AtlasCloudfastest | 51 |
| Phala | 27 |
| Venice | 21 |
| DeepInfrafastest | 18 |
| SiliconFlow | 11 |
| Alibaba | 10 |
| Novita | 8 |
| Darkbloom | 6 |
| qwenfastest | — |
| Artificial Analysis | 0 |