Qwen: Qwen3.5-Flash
qwen/qwen3.5-flash-02-23
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...
Prompt $/1M
$0.065
Completion $/1M
$0.260
Context
1,000,000 tok
Best provider
77 tok/s · Alibaba
Specs
- Model id
- qwen/qwen3.5-flash-02-23
- Display name
- Qwen: Qwen3.5-Flash
- Modality
- text+image+video->text
- Tokenizer
- Qwen3
- Instruct type
- —
- Context length
- 1,000,000
- Max completion
- 65,536
- Prompt $/1M
- $0.065
- Completion $/1M
- $0.260
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionyes
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Alibabafastest | 77 |