Benchmarks

NVIDIA: Nemotron 3 Nano 30B A3B

nvidia/nemotron-3-nano-30b-a3b

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

Prompt $/1M
$0.050
Completion $/1M
$0.200
Context
262,144 tok
Best provider
208 tok/s · Novita

Specs

Model id
nvidia/nemotron-3-nano-30b-a3b
Display name
NVIDIA: Nemotron 3 Nano 30B A3B
Modality
text->text
Tokenizer
Other
Instruct type
Context length
262,144
Max completion
228,000
Prompt $/1M
$0.050
Completion $/1M
$0.200
Free tier
no

Capabilities

  • Function callingno
  • JSON modeno
  • System promptno
  • Visionno
  • Moderatedno

Benchmark scores

Each axis is one benchmark, normalized to the best model in the field.

No benchmark data for this model.

Speed by provider

Providertok/s
Novitafastest208
Ambientfastest121
DeepInfra74
Nebius39
NVIDIA: Nemotron 3 Nano 30B A3B - Benchmarks, Pricing and Speed | TryAii