Qwen: Qwen3.5-35B-A3B
qwen/qwen3.5-35b-a3b
The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...
Prompt $/1M
$0.140
Completion $/1M
$1.000
Context
262,144 tok
Best provider
209 tok/s · WandB
Specs
- Model id
- qwen/qwen3.5-35b-a3b
- Display name
- Qwen: Qwen3.5-35B-A3B
- Modality
- text+image+video->text
- Tokenizer
- Qwen3
- Instruct type
- —
- Context length
- 262,144
- Max completion
- 262,144
- Prompt $/1M
- $0.140
- Completion $/1M
- $1.000
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionyes
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| WandBfastest | 209 |
| Artificial Analysis | 167 |
| DeepInfra | 138 |
| AkashML | 135 |
| Parasail | 111 |
| Venice | 109 |
| Alibaba | 100 |
| Ambient | 93 |
| DekaLLM | 79 |
| AtlasCloud | 65 |
| NextBit | 47 |
| SiliconFlow | 23 |