Benchmarks

NVIDIA: Nemotron Nano 9B V2

nvidia/nemotron-nano-9b-v2

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

Prompt $/1M
$0.040
Completion $/1M
$0.160
Context
131,072 tok
Best provider
14 tok/s · DeepInfra

Specs

Model id
nvidia/nemotron-nano-9b-v2
Display name
NVIDIA: Nemotron Nano 9B V2
Modality
text->text
Tokenizer
Other
Instruct type
Context length
131,072
Max completion
16,384
Prompt $/1M
$0.040
Completion $/1M
$0.160
Free tier
no

Capabilities

  • Function callingno
  • JSON modeno
  • System promptno
  • Visionno
  • Moderatedno

Benchmark scores

Each axis is one benchmark, normalized to the best model in the field.

No benchmark data for this model.

Speed by provider

Providertok/s
DeepInfrafastest14
NVIDIA: Nemotron Nano 9B V2 - Benchmarks, Pricing and Speed | TryAii