NVIDIA: Nemotron Nano 12B 2 VL
nvidia/nemotron-nano-12b-v2-vl
NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...
Prompt $/1M
$0.200
Completion $/1M
$0.600
Context
131,072 tok
Best provider
116 tok/s · DeepInfra
Specs
- Model id
- nvidia/nemotron-nano-12b-v2-vl
- Display name
- NVIDIA: Nemotron Nano 12B 2 VL
- Modality
- text+image+video->text
- Tokenizer
- Other
- Instruct type
- —
- Context length
- 131,072
- Max completion
- 16,384
- Prompt $/1M
- $0.200
- Completion $/1M
- $0.600
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionyes
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| DeepInfrafastest | 116 |
| nvidiafastest | — |