Qwen: Qwen3 30B A3B Instruct 2507
qwen/qwen3-30b-a3b-instruct-2507
Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...
Prompt $/1M
$0.100
Completion $/1M
$0.300
Context
262,144 tok
Best provider
89 tok/s · Phala
Specs
- Model id
- qwen/qwen3-30b-a3b-instruct-2507
- Display name
- Qwen: Qwen3 30B A3B Instruct 2507
- Modality
- text->text
- Tokenizer
- Qwen3
- Instruct type
- —
- Context length
- 262,144
- Max completion
- —
- Prompt $/1M
- $0.100
- Completion $/1M
- $0.300
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Phalafastest | 89 |
| Alibaba | 87 |
| WandB | 67 |
| Nebius | 59 |
| DekaLLM | 47 |
| StreamLake | 32 |
| SiliconFlow | 20 |
| Venice | 16 |
| AtlasCloud | 11 |