Benchmarks

NVIDIA: Nemotron Nano 12B 2 VL

nvidia/nemotron-nano-12b-v2-vl

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...

Prompt $/1M
$0.200
Completion $/1M
$0.600
Context
131,072 tok
Best provider
116 tok/s · DeepInfra

Specs

Model id
nvidia/nemotron-nano-12b-v2-vl
Display name
NVIDIA: Nemotron Nano 12B 2 VL
Modality
text+image+video->text
Tokenizer
Other
Instruct type
Context length
131,072
Max completion
16,384
Prompt $/1M
$0.200
Completion $/1M
$0.600
Free tier
no

Capabilities

  • Function callingno
  • JSON modeno
  • System promptno
  • Visionyes
  • Moderatedno

Benchmark scores

Each axis is one benchmark, normalized to the best model in the field.

No benchmark data for this model.

Speed by provider

Providertok/s
DeepInfrafastest116
nvidiafastest
NVIDIA: Nemotron Nano 12B 2 VL - Benchmarks, Pricing and Speed | TryAii