Benchmarks

NVIDIA: Nemotron 3.5 Lightning

nvidia/nemotron-3.5-lightning

Nemotron 3.5 Lightning (NVIDIA) is a low-cost open-weight model priced near the bottom of the board, at $0.08/$0.2 per million input/output tokens, served by 3 independent hosts. For the money it delivers #158 of 248 on the TryAii Score. It's among the fastest models we track, at roughly 24-161 tokens/sec.

Prompt $/1M
$0.080
Completion $/1M
$0.200
Context
262,144 tok
Best provider
264 tok/s · Artificial Analysis

Specs

Model id
nvidia/nemotron-3.5-lightning
Display name
NVIDIA: Nemotron 3.5 Lightning
Modality
text->text
Tokenizer
Other
Instruct type
—
Context length
262,144
Max completion
131,072
Prompt $/1M
$0.080
Completion $/1M
$0.200
Free tier
no

Capabilities

  • Function callingno
  • JSON modeno
  • System promptno
  • Visionno
  • Moderatedno

Benchmark scores

Each axis is one benchmark, normalized to the best model in the field.

No benchmark data for this model.

Speed by provider

Providertok/s
Artificial Analysis264
CoreWeavefastest126
DeepInfra31
Venice5
nvidiafastest—