Benchmarks

NVIDIA: Nemotron 3.5 Lightning

nvidia/nemotron-3.5-lightning

Nemotron 3.5 Lightning (NVIDIA) is a low-cost open-weight model priced near the bottom of the board, at $0.1/$0.25 per million input/output tokens, served by 2 independent hosts. For the money it delivers #132 of 222 on the TryAii Score. It's among the fastest models we track, at roughly 104-270 tokens/sec.

Prompt $/1M
$0.100
Completion $/1M
$0.250
Context
262,144 tok
Best provider
354 tok/s · DeepInfra

Specs

Model id
nvidia/nemotron-3.5-lightning
Display name
NVIDIA: Nemotron 3.5 Lightning
Modality
text->text
Tokenizer
Other
Instruct type
Context length
262,144
Max completion
262,144
Prompt $/1M
$0.100
Completion $/1M
$0.250
Free tier
no

Capabilities

  • Function callingno
  • JSON modeno
  • System promptno
  • Visionno
  • Moderatedno

Benchmark scores

Each axis is one benchmark, normalized to the best model in the field.

No benchmark data for this model.

Speed by provider

Providertok/s
DeepInfrafastest354
CoreWeave20
NVIDIA: Nemotron 3.5 Lightning - Benchmarks, Pricing and Speed | TryAii