Qwen: Qwen2.5 7B Instruct
qwen/qwen-2.5-7b-instruct
Qwen2.5 7B Instruct (Qwen) is a compact 7B open-weight model priced near the bottom of the board, at $0.1/$0.2 per million input/output tokens, served by 3 independent hosts. For the money it delivers #220 of 234 on the TryAii Score. Its best showing is #7 on DROP.
Prompt $/1M
$0.100
Completion $/1M
$0.200
Context
32,768 tok
Best provider
56 tok/s · Together
Specs
- Model id
- qwen/qwen-2.5-7b-instruct
- Display name
- Qwen: Qwen2.5 7B Instruct
- Modality
- text->text
- Tokenizer
- Qwen
- Instruct type
- chatml
- Context length
- 32,768
- Max completion
- 29,491
- Prompt $/1M
- $0.100
- Completion $/1M
- $0.200
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptyes
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Together | 56 |
| AtlasCloud | 34 |
| Phalafastest | 23 |