Qwen: Qwen3 8B
qwen/qwen3-8b
Qwen3 8B (Qwen) is a compact 8B open-weight model priced near the bottom of the board, at $0.117/$0.455 per million input/output tokens, served by 1 independent host. For the money it delivers #198 of 235 on the TryAii Score. It's reasonably quick at roughly 49-51 tokens/sec.
Prompt $/1M
$0.117
Completion $/1M
$0.455
Context
131,072 tok
Best provider
52 tok/s · AtlasCloud
Specs
- Model id
- qwen/qwen3-8b
- Display name
- Qwen: Qwen3 8B
- Modality
- text->text
- Tokenizer
- Qwen3
- Instruct type
- qwen3
- Context length
- 131,072
- Max completion
- 8,192
- Prompt $/1M
- $0.117
- Completion $/1M
- $0.455
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptyes
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| AtlasCloudfastest | 52 |
| Alibabafastest | 48 |
| qwenfastest | — |
| Artificial Analysis | 0 |