Qwen: Qwen3 235B A22B Thinking 2507
qwen/qwen3-235b-a22b-thinking-2507
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...
Prompt $/1M
$0.300
Completion $/1M
$3.000
Context
262,144 tok
Best provider
73 tok/s · WandB
Specs
- Model id
- qwen/qwen3-235b-a22b-thinking-2507
- Display name
- Qwen: Qwen3 235B A22B Thinking 2507
- Modality
- text->text
- Tokenizer
- Qwen3
- Instruct type
- qwen3
- Context length
- 262,144
- Max completion
- 32,768
- Prompt $/1M
- $0.300
- Completion $/1M
- $3.000
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptyes
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| WandBfastest | 73 |
| Alibabafastest | 61 |
| Venicefastest | 48 |
| DeepInfra | 47 |
| AtlasCloud | 46 |
| SiliconFlow | 25 |
| Novita | 23 |