Qwen: Qwen3 Next 80B A3B Thinking
qwen/qwen3-next-80b-a3b-thinking
Qwen3 Next 80B A3B Thinking is a 80B open-weight model (3B active per token), ranking #166 of 235 on the TryAii Score. It's served by 3 independent hosts from $0.15/$1.2 per million input/output tokens. It's among the fastest models we track, at roughly 93-184 tokens/sec.
Prompt $/1M
$0.150
Completion $/1M
$1.200
Context
262,144 tok
Best provider
201 tok/s · Alibaba
Specs
- Model id
- qwen/qwen3-next-80b-a3b-thinking
- Display name
- Qwen: Qwen3 Next 80B A3B Thinking
- Modality
- text->text
- Tokenizer
- Qwen3
- Instruct type
- —
- Context length
- 262,144
- Max completion
- 32,768
- Prompt $/1M
- $0.150
- Completion $/1M
- $1.200
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Alibabafastest | 201 |
| Novitafastest | 185 |
| AtlasCloud | 171 |
| Nebius | 93 |
| 67 | |
| qwenfastest | — |