Qwen: Qwen3 235B A22B Instruct 2507
qwen/qwen3-235b-a22b-2507
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...
Prompt $/1M
$0.090
Completion $/1M
$0.550
Context
262,144 tok
Best provider
78 tok/s · WandB
Specs
- Model id
- qwen/qwen3-235b-a22b-2507
- Display name
- Qwen: Qwen3 235B A22B Instruct 2507
- Modality
- text->text
- Tokenizer
- Qwen3
- Instruct type
- —
- Context length
- 262,144
- Max completion
- 16,384
- Prompt $/1M
- $0.090
- Completion $/1M
- $0.550
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| WandBfastest | 78 |
| Alibaba | 47 |
| Cerebras | 46 |
| 41 | |
| Nebius | 40 |
| StreamLake | 37 |
| AtlasCloud | 35 |
| Together | 35 |
| Parasail | 27 |
| Novita | 24 |
| DeepInfra | 19 |
| Venice | 17 |
| Friendli | 13 |
| Chutes | 8 |
| SiliconFlow | 8 |
| GMICloud | 7 |