Qwen: Qwen3 Next 80B A3B Thinking
qwen/qwen3-next-80b-a3b-thinking
Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...
Prompt $/1M
$0.098
Completion $/1M
$0.780
Context
262,144 tok
Best provider
185 tok/s · Novita
Specs
- Model id
- qwen/qwen3-next-80b-a3b-thinking
- Display name
- Qwen: Qwen3 Next 80B A3B Thinking
- Modality
- text->text
- Tokenizer
- Qwen3
- Instruct type
- —
- Context length
- 262,144
- Max completion
- 32,768
- Prompt $/1M
- $0.098
- Completion $/1M
- $0.780
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Novitafastest | 185 |
| AtlasCloud | 171 |
| Alibabafastest | 151 |
| 136 | |
| Nebius | 29 |
| qwenfastest | — |