NVIDIA: Nemotron 3 Nano 30B A3B
nvidia/nemotron-3-nano-30b-a3b
Nemotron 3 Nano 30B A3B is a NVIDIA model that isn't ranked on the TryAii Score yet — only 2 benchmarks on file so far. It lists at $0.05/$0.2 per million input/output tokens with a 262K-token context. It's among the fastest models we track, at roughly 121-162 tokens/sec.
Prompt $/1M
$0.050
Completion $/1M
$0.200
Context
262,144 tok
Best provider
201 tok/s · Crusoe
Specs
- Model id
- nvidia/nemotron-3-nano-30b-a3b
- Display name
- NVIDIA: Nemotron 3 Nano 30B A3B
- Modality
- text->text
- Tokenizer
- Other
- Instruct type
- —
- Context length
- 262,144
- Max completion
- 235,929
- Prompt $/1M
- $0.050
- Completion $/1M
- $0.200
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Crusoefastest | 201 |
| Novita | 163 |
| Nebius | 135 |
| Ambientfastest | 121 |
| DeepInfra | 67 |
| nvidiafastest | — |