DeepSeek: DeepSeek V4 Flash
deepseek/deepseek-v4-flash
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
Prompt $/1M
$0.094
Completion $/1M
$0.188
Context
1,048,576 tok
Best provider
113 tok/s · Artificial Analysis
Specs
- Model id
- deepseek/deepseek-v4-flash
- Display name
- DeepSeek: DeepSeek V4 Flash
- Modality
- text->text
- Tokenizer
- DeepSeek
- Instruct type
- —
- Context length
- 1,048,576
- Max completion
- —
- Prompt $/1M
- $0.094
- Completion $/1M
- $0.188
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Artificial Analysis | 113 |
| DekaLLM | 79 |
| NextBitfastest | 78 |
| Alibabafastest | 69 |
| DeepSeek | 68 |
| Baidu | 68 |
| AtlasCloud | 55 |
| StreamLake | 54 |
| Fireworks | 47 |
| Novita | 45 |
| GMICloud | 38 |
| Parasail | 37 |
| WandB | 34 |
| Venice | 30 |
| Cloudflare | 29 |
| Morph | 27 |
| DeepInfra | 24 |
| Wafer | 18 |
| Ambient | 15 |
| SiliconFlow | 14 |
| AkashML | 12 |
| DigitalOcean | 8 |
| Io Net | 6 |