NVIDIA: Nemotron Nano 9B V2
nvidia/nemotron-nano-9b-v2
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...
Prompt $/1M
$0.040
Completion $/1M
$0.160
Context
131,072 tok
Best provider
14 tok/s · DeepInfra
Specs
- Model id
- nvidia/nemotron-nano-9b-v2
- Display name
- NVIDIA: Nemotron Nano 9B V2
- Modality
- text->text
- Tokenizer
- Other
- Instruct type
- —
- Context length
- 131,072
- Max completion
- 16,384
- Prompt $/1M
- $0.040
- Completion $/1M
- $0.160
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| DeepInfrafastest | 14 |