Qwen: Qwen3.5 397B A17B
qwen/qwen3.5-397b-a17b
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
Prompt $/1M
$0.390
Completion $/1M
$2.340
Context
262,144 tok
Best provider
158 tok/s · Together
Specs
- Model id
- qwen/qwen3.5-397b-a17b
- Display name
- Qwen: Qwen3.5 397B A17B
- Modality
- text+image+video->text
- Tokenizer
- Qwen3
- Instruct type
- —
- Context length
- 262,144
- Max completion
- 65,536
- Prompt $/1M
- $0.390
- Completion $/1M
- $2.340
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionyes
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Togetherfastest | 158 |
| Wafer | 99 |
| Nebius | 67 |
| Artificial Analysis | 65 |
| Morph | 54 |
| AtlasCloudfastest | 54 |
| GMICloud | 48 |
| Parasail | 44 |
| DeepInfra | 30 |
| Venice | 29 |
| StreamLake | 25 |
| Phala | 17 |
| Novita | 17 |
| Alibaba | 16 |
| Chutes | 12 |
| DigitalOcean | 7 |