NVIDIA: Nemotron 3.5 Lightning
nvidia/nemotron-3.5-lightning
Nemotron 3.5 Lightning (NVIDIA) is a low-cost open-weight model priced near the bottom of the board, at $0.08/$0.2 per million input/output tokens, served by 3 independent hosts. For the money it delivers #158 of 248 on the TryAii Score. It's among the fastest models we track, at roughly 24-161 tokens/sec.
Prompt $/1M
$0.080
Completion $/1M
$0.200
Context
262,144 tok
Best provider
264 tok/s · Artificial Analysis
Specs
- Model id
- nvidia/nemotron-3.5-lightning
- Display name
- NVIDIA: Nemotron 3.5 Lightning
- Modality
- text->text
- Tokenizer
- Other
- Instruct type
- —
- Context length
- 262,144
- Max completion
- 131,072
- Prompt $/1M
- $0.080
- Completion $/1M
- $0.200
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Artificial Analysis | 264 |
| CoreWeavefastest | 126 |
| DeepInfra | 31 |
| Venice | 5 |
| nvidiafastest | — |