Qwen: Qwen3 30B A3B Instruct 2507
qwen/qwen3-30b-a3b-instruct-2507
Qwen3 30B A3B Instruct 2507 is a Qwen model that isn't ranked on the TryAii Score yet — only 2 benchmarks on file so far. It lists at $0.0481/$0.193 per million input/output tokens with a 262K-token context. It's reasonably quick at roughly 20-60 tokens/sec.
Prompt $/1M
$0.048
Completion $/1M
$0.193
Context
262,144 tok
Best provider
70 tok/s · Phala
Specs
- Model id
- qwen/qwen3-30b-a3b-instruct-2507
- Display name
- Qwen: Qwen3 30B A3B Instruct 2507
- Modality
- text->text
- Tokenizer
- Qwen3
- Instruct type
- —
- Context length
- 262,144
- Max completion
- 32,000
- Prompt $/1M
- $0.048
- Completion $/1M
- $0.193
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Phala | 70 |
| DekaLLM | 64 |
| WandB | 63 |
| CoreWeavefastest | 53 |
| Nebius | 47 |
| Alibaba | 42 |
| StreamLake | 30 |
| Venice | 16 |
| SiliconFlow | 14 |
| AtlasCloud | 11 |