Qwen: Qwen3 32B
qwen/qwen3-32b
Qwen3 32B is a 32B open-weight model, ranking #176 of 235 on the TryAii Score. Its best showing is #16 on MATH. It's served by 7 independent hosts from $0.08/$0.28 per million input/output tokens — about a third of what models at its level usually charge.
Prompt $/1M
$0.080
Completion $/1M
$0.280
Context
131,072 tok
Best provider
210 tok/s · Groq
Specs
- Model id
- qwen/qwen3-32b
- Display name
- Qwen: Qwen3 32B
- Modality
- text->text
- Tokenizer
- Qwen3
- Instruct type
- qwen3
- Context length
- 131,072
- Max completion
- 16,384
- Prompt $/1M
- $0.080
- Completion $/1M
- $0.280
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptyes
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Groqfastest | 210 |
| Alibaba | 78 |
| Novita | 71 |
| AtlasCloud | 23 |
| DeepInfrafastest | 21 |
| Nebius | 17 |
| SiliconFlow | 13 |
| Chutes | 9 |
| qwenfastest | — |
| Artificial Analysis | 0 |